3 ms·
A) we plan on thoroughly testing that with Mind2Web dataset. They have a very robust set of (persistant) selectors B) so, shadcn for prompts for web agents hah
by gregpr07 2y ago
A) we plan on thoroughly testing that with Mind2Web dataset. They have a very robust set of (persistant) selectors
B) so, shadcn for prompts for web agents haha :) but I agree, that would be SICK! Just go to browseruse and get the prompt for your specific use case
- maggreenWAI 2y agoA) For Mind2Web: because there are multiple ways to reach a goal state - any thoughts how to evaluate if a task was successful? Should we let the LLM/ other LLM evaluate it?