4 ms·
After retrieving all the text data from the webpage (using a headless browser), GPT is used to filter out all the noise and extract the actual information reque
by semanser 3y ago
After retrieving all the text data from the webpage (using a headless browser), GPT is used to filter out all the noise and extract the actual information requested in the request schema. Let's say you request for {"product_name": "string"}. GPT will retrieve that product name from the webpage and return the correctly formatted JSON with the fields you requested.
It works pretty similarly to GraphQL when you define a schema that you want, and the backend returns the exact data that you requested. But in this case, the data is received from the webpage that you provided.
- KomoD 3y agoCan the results really be trusted then? Isn't it possible for GPT to make something up that doesn't exist on the site
- deleted 3y ago[deleted]