3 ms·
Out of curiosity, what did you need beyond something like OpenCV/Metal + Tesseract?
by staticautomatic 9y ago
Out of curiosity, what did you need beyond something like OpenCV/Metal + Tesseract?
- pault 9y agoThe client needed an OCR solution for supplier invoices with a variety of layouts and a combination of printed and hand-written characters, and didn't have the budget for a bespoke solution. To be fair, it's a very hard problem, I was just surprised that given all the much hyped recent advancements in deep learning for computer vision, most of the solutions in the market seem to be running on decades old technology.
- ocrcustomserver 9y agoWell it is more complex than it appears. Extracting data from documents requires a solution which uses OCR but is a different product (e.g. ABBYY FlexiCapture). This is most commonly referred to as zonal OCR and comes with the added functionality of handling multiple templates, defining zones/fields, specifying special rules for fields, verification process for manual inspection (e.g. triggered when the image receives a low confidence recognition score) etc. This is different and more complex than a product that does full page OCR (e.g. ABBYY Finereader). Handwritten OCR is a whole different story. The products that do zonal OCR will fail to recognize handwritten text, unless it's in boxes (PDF forms). I'm working on a prototype that can handle handwritten text outside of boxes too.