Comments (2)
Only PDF input is supported for now.
img2pdf does a good job at converting most JPEGs and PNGs to PDF.
And yes, that's a really crappy error message, which I should fix.
On Tue, 12 Jan 2016 at 15:06 Shaun [email protected] wrote:
When I try to run:
sudo ocrmypdf --verbose 3 eiffeljpg eiffelpdf
I get:
Original exception:
Exception #1
'builtinsTypeError(Can't convert 'list' object to str implicitly)' raised in
Task = def ocrmypdfmainsplit_pages():
Job = [[] -> /comgithubocrmypdf45n_qza7/*pagepdf, , [], <_threadlock>]Traceback (most recent call last):
File "/usr/local/lib/python34/dist-packages/ruffus/taskpy", line 751, in run_pooled_job_without_exceptions
register_cleanup, touch_files_only)
File "/usr/local/lib/python34/dist-packages/ruffus/taskpy", line 567, in job_wrapper_io_files
ret_val = user_defined_work_func(_params)
File "/usr/local/lib/python34/dist-packages/ocrmypdf/mainpy", line 415, in split_pages
npages = qpdfget_npages(input_file)
File "/usr/local/lib/python34/dist-packages/ocrmypdf/qpdfpy", line 68, in get_npages
universal_newlines=True, close_fds=True)
File "/usr/lib/python34/subprocesspy", line 607, in check_output
with Popen(_popenargs, stdout=PIPE, **kwargs) as process:
File "/usr/lib/python34/subprocesspy", line 859, in init
restore_signals, start_new_session)
File "/usr/lib/python34/subprocesspy", line 1395, in _execute_child
restore_signals, start_new_session, preexec_fn)
TypeError: Can't convert 'list' object to str implicitlyIf I try the same thing on a PDF file it works fine Thanks!
—
Reply to this email directly or view it on GitHub
#42.
from ocrmypdf.
Next release fixes the error message.
By the way, there should be no need to sudo ocrmypdf
. You don't have to trust me with root access to your system.
from ocrmypdf.
Related Issues (20)
- [Bug]: crashes with tesseract 5.4.0 HOT 8
- [Bug]: ocrmypdf 16.3.1 fails on a file on Arch that 13.4.0 on Ubuntu handles well HOT 1
- [Feature]: Alternative AI OCR "surya" as opposed to EasyOCR, Just found it today and it dominated the accuracy and speed of Tesseract & EasyOCR HOT 3
- [Bug]: Paperless-ngx Release 2.9.0 Ghostscript rasterizing failed HOT 1
- [Bug]: MetadataProgress does not respect progress_bar=False argument
- [Bug]: No errors and no output for large DPI files HOT 2
- [Bug]: `lots of diacritics - possibly poor OCR` but using standalone tesseract works perfectly HOT 1
- [Bug]: ocrmypdf (16.3.1) and Tesseract 5.4.1 HOT 3
- [Bug]: Existing text is completely replaced with other characters HOT 3
- [Request]: Please make rich logging library an optional dependency HOT 1
- [Feature]: Enable execution on GPU HOT 1
- [Bug]: doesn't always parse Latin with diacritics HOT 3
- Output file images are corrupted HOT 1
- [Bug]: OSError: [Errno 28] No space left on device HOT 4
- [Bug]: problem with tif "DPI is not credible". Estimate dpi HOT 3
- [Bug]: Ghostscript can't create a PDF/A-file (Page object was reserved for an Annotation destination) HOT 3
- [Bug]: KeyError: '/Subtype'
- [Bug]: Ghostscript rasterizing failed HOT 3
- [Bug]: files signed with a-trust are not recognised as digitally signed and hence processed HOT 1
- --sidecar writes text content and messages to file HOT 2
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from ocrmypdf.