3 ms·
Not sure what your budget is but I’ve used AWS for handling PDFs and it’s been pretty good at detecting content via boundary boxes.
by ensemblehq 3y ago
Not sure what your budget is but I’ve used AWS for handling PDFs and it’s been pretty good at detecting content via boundary boxes.
- anamexis 3y agoAWS as in Amazon Web Services? And if so, can you be more specific?
- navanchauhan 3y agoHere are some more options: * AWS Textract [0] * Microsoft Azure [1] * Google Cloud Vision [2] I personally use Azure, combined with OCR correction using GPT to convert a scan of my daily journal (Apple Notes creates a PDF that is nothing but a bunch of images) -> Markdown -> Extract tasks and then add them to my Reminders app using CalDav. Azure has one of the best OCR for handwritten text, but for normal document extraction (read: printed text), any service would do a reasonable job. [0] https://docs.aws.amazon.com/prescriptive-guidance/latest/patterns/automatically-extract-content-from-pdf-files-using-amazon-textract.html https://docs.aws.amazon.com/prescriptive-guidance/latest/pat... [1] https://learn.microsoft.com/en-us/azure/data-factory/solution-template-extract-data-from-pdf https://learn.microsoft.com/en-us/azure/data-factory/solutio... [2] https://cloud.google.com/vision/docs/pdf https://cloud.google.com/vision/docs/pdf
- swsieber 3y agoCould you explain how you use GPT for ocr correction?
- navanchauhan 3y agohttps://promptbase.com/prompt/ocr-text-fixer https://promptbase.com/prompt/ocr-text-fixer This is the prompt I bought from promptbase. You basically provide GPT with some examples on possible OCR errors, and then you give it the OCRed text and it tries to correct it
- swsieber 3y agoFascinating. Thanks!