4 ms·
Nope, don't have it yet. Would be really cool to plop in a PDF that's made up of just images, and tell it to describe each page of the PDF to me. As a blind pe
by devinprater 3y ago
Nope, don't have it yet. Would be really cool to plop in a PDF that's made up of just images, and tell it to describe each page of the PDF to me. As a blind person, that'd just... Be a dream come true.
- kridsdale3 3y agoI'm very excited on your behalf for what is about to happen.
- cco 3y agoYou can directly upload images, both on the web and mobile. It works really well. In fact I've used both images and voice to describe things and it works like a charm. You should already have that if you pay for plus.
- spdustin 3y agoSadly, I don't think that'll work. They use the same headless browser setup used by Browse with Bing, and it only extracts the baked-in text from a PDF.
- igemal 3y agoRight now I'm working on a fork of a little web app that parses a resume and spits it into JSON format with GPT (I'm working on stuff like OCR for a scanned pdf). https://github.com/IsaacGemal/nlp-resume-parser https://github.com/IsaacGemal/nlp-resume-parser I feel like it wouldn't be that difficult to fork it again, but rewrite the main function so it sends the pdf to some sort of GPT-Vision, and write the output again with a regular GPT api call. Does such a tool not exist? Or maybe I have to wait for image support via the API.