3 ms·
I would check out vision models as a technique to go around OCR errors. ColPali is the standard implementation & SOTA. Much better than OCR. We maintain a read
by jonathan-adly 2y ago
I would check out vision models as a technique to go around OCR errors.
ColPali is the standard implementation & SOTA. Much better than OCR. We maintain a ready to go retrieval API that implements this: https://github.com/tjmlabs/ColiVara https://github.com/tjmlabs/ColiVara