fritz-hh 0c46a723bd OCRmyPDF.sh: many improvements!
- automatic analysis of jhove validation report
- quiet generation of PDF/A with gs
- deletion of tmp files
- Corrected issue that lead to crash at page 8
- Improved log
2013-04-18 23:13:06 +02:00
2013-04-18 23:13:06 +02:00
2013-04-18 10:44:10 +02:00

OCRmyPDF

Collection of script aimed at generating searchable PDF files from PDF files containing only images

ATTENTION: The scripts are still in development phase, please do not use!!!!

Install

TODO

Install java: cd /usr/ports/java/openjdk7/ && make install clean

Install jhove: download jhove from here: http://sourceforge.net/projects/jhove/files/jhove/ After extracting the JHOVE files to some directory "jhove", you have to edit the file "jhove/conf/jhove.conf" and change something in "something" to the actual directory (ending in "/jhove").

S
Description
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
http://ocrmypdf.readthedocs.io/ Readme MPL-2.0
80 MiB
Languages
Python 97.9%
Shell 1.8%
Dockerfile 0.3%