0c46a723bd930ed32735da06368b97e78be7ac00
- automatic analysis of jhove validation report - quiet generation of PDF/A with gs - deletion of tmp files - Corrected issue that lead to crash at page 8 - Improved log
OCRmyPDF
Collection of script aimed at generating searchable PDF files from PDF files containing only images
ATTENTION: The scripts are still in development phase, please do not use!!!!
Install
TODO
Install java: cd /usr/ports/java/openjdk7/ && make install clean
Install jhove: download jhove from here: http://sourceforge.net/projects/jhove/files/jhove/ After extracting the JHOVE files to some directory "jhove", you have to edit the file "jhove/conf/jhove.conf" and change something in "something" to the actual directory (ending in "/jhove").
Description
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
http://ocrmypdf.readthedocs.io/
Readme
MPL-2.0
80 MiB
Languages
Python
97.9%
Shell
1.8%
Dockerfile
0.3%