4 ms·
It's a really good model from my testing so far. You can see the difference in how it tries to use tools to the greatest extent when answering a question, espec
by joshmlewis 1y ago
It's a really good model from my testing so far. You can see the difference in how it tries to use tools to the greatest extent when answering a question, especially compared to 4.1 and o3. In this example it used 6! tool calls in the first response to try and collect as much info as possible.
https://promptslice.com/share/b-2ap_rfjeJgIQsG https://promptslice.com/share/b-2ap_rfjeJgIQsG
- hollownobody 1y ago720 tool calls? Amazing!
- joshmlewis 1y agoWhere'd you get 720 from?
- terhechte 1y agothe _6!_
- brian626 1y agoMath pun… 6! = Factorial(6) = 720
- joshmlewis 1y agoWhoosh, it went right over my head.
- Zone3513 1y agoThat movie doesn't even exist. There is no Thunder Run from 2025.
- joshmlewis 1y agoThe data is made up, the point is to see how models respond to the same input / scenario. You're able to create whatever tools you want and import real data or it'll generate fake tool responses for you based on the prompt and tool definition. Disclaimer: I made PromptSlice for creating and comparing prompts, tools, and models.
- mustaphah 1y agoIs there any value in using XML elements to guide the model instead of simple text (e.g., "Recommendation criteria:")?
- flexagoon 1y agoXML tags generally help models understand prompts better. That's how most official system prompts are written and what the Anthropic prompting guide says.