4 ms·
Hi! I'm a project dev. In case it's helpful, there's lots of Q&A in this previous thread: https://news.ycombinator.com/item?id=18330876 https://news.ycombinator
by JackC 8y ago
Hi! I'm a project dev. In case it's helpful, there's lots of Q&A in this previous thread: https://news.ycombinator.com/item?id=18330876 https://news.ycombinator.com/item?id=18330876
I'll try to answer any questions that pop up here as well.
- ocrcustomserver 8y agoWhich OCR did you use?
- JackC 8y agoABBYY FineReader -- I don't have the specific version in front of me unfortunately. In addition to the structured text we're currently serving through the API, we also have 300DPI color scans and per-word coordinates and confidence scores, so there's a lot more we can do with the OCR data that isn't exposed yet.
- just_myles 8y agoFor those docs that require more finesse (This is OCR :D), was there a massive manual effort involved?
- just_myles 8y agoNice question. That's what I want to know.
- deleted 8y ago[deleted]