5 ms·
that's difficult as well, how do you k ow where to split?
by throwaw12 2mo ago
that's difficult as well, how do you k ow where to split?
- kgwgk 2mo agoText is often written as separate lines (and paragraphs) at least in some languages.
- vrganj 2mo agoPresumably a small cheap model could do that part?
- grog454 2mo agoOverlap the splits?
- wongarsu 2mo agoLet the model do the splitting. A 800x800px image should be enough to make those decisions
- johndough 2mo agoThere are models specifically for splitting an image into text regions, e.g. PP-DocLayoutV3 https://huggingface.co/PaddlePaddle/PP-DocLayoutV3 https://huggingface.co/PaddlePaddle/PP-DocLayoutV3 I am using a stripped-down minimal version of it which I uploaded here, since I am not a fan of huge dependency trees: https://github.com/99991/simple-pp-doclayoutv3 https://github.com/99991/simple-pp-doclayoutv3 Another recent model for this task is Unlimited-OCR: https://github.com/baidu/Unlimited-OCR https://github.com/baidu/Unlimited-OCR