4 ms·
Nice stuff! I found an error in the chinese demo, with the example you provided (4th character wasn't the same). I know no OCR is perfect, but IMHO at least yo
by zikero 5y ago
Nice stuff!
I found an error in the chinese demo, with the example you provided (4th character wasn't the same). I know no OCR is perfect, but IMHO at least your own demo should be free of errors.
- mkl 5y agoThere's one in the English demo too: "hail!" -> "haill". They're both pretty bad images though. In practice I've found (command line) Tesseract very accurate on 300dpi scans of printed documents, with colour/greyscale, not binary.
- mdp2021 5y ago> at least your own demo should be free of errors :) That would be a dishonest demo. You try to show how well it works, not that it works perfectly well (which is false). Edit: especially since we know that OCR is hardly perfect - we expect errors to be minimized, not absent, and the first interest is to see where the engine fails.