4 ms·
For instace Llamaparse(https://docs.llamaindex.ai/en/stable/llama_cloud/llama_parse/ https://docs.llamaindex.ai/en/stable/llama_cloud/llama_parse...)uses LLMs f
by constantinum 2y ago
For instace Llamaparse(https://docs.llamaindex.ai/en/stable/llama_cloud/llama_parse/ https://docs.llamaindex.ai/en/stable/llama_cloud/llama_parse...)uses LLMs for pdf text extraction, but the problem is hallucination. e.g > https://github.com/run-llama/llama_parse/issues/420 https://github.com/run-llama/llama_parse/issues/420
There is also LLMWhisperer that preserves the layout(tables, checkboxes, forms)and hence the context. https://pg.llmwhisperer.unstract.com/ https://pg.llmwhisperer.unstract.com/
- cpursley 2y agoIs this open source? Is it slow Python? That's where I'm stuck.
- constantinum 2y agoThis is not open-source. It has high accuracy and it is faster too. All you need is to point your documents to the API.