2 ms·
PDF Hell: Why is extracting data still a nightmare?
- chrisjj 8mo ago> Text in a PDF file is ... lacks any logical or semantic structure. Checks "PDF". Checks "lack". Hmm.
- ftchd 8mo agoI found that Claude Sonnet 4.6 solves all of this very easily No workflows and 0 setup, you just give it the PDF with photos of text and it spits out a perfect `docx` (not just in english)