3 ms·
I use https://kebekus.gitlab.io/scantools https://kebekus.gitlab.io/scantools for scanning, it builds on top of tesseract and works great for pdf enhancements
by ce4 4y ago
I use https://kebekus.gitlab.io/scantools https://kebekus.gitlab.io/scantools for scanning, it builds on top of tesseract and works great for pdf enhancements
- rjzzleep 4y agoYou might be interested in https://github.com/ocrmypdf/OCRmyPDF https://github.com/ocrmypdf/OCRmyPDF then. It does quite some preprocessing on the PDF pages before passing it on to tesseract.
- angrygoat 4y agoI've found ocrmypdf to be excellent: the only issue I've had is with PDFs with differing page sizes; it seems to scale everything up to the size of the largest page, which can be a bit of a pain.