4 ms·
It is on top for many benchmarks, only not the coding/agentic ones. Still one of the most intelligent models overall, most likely to get any question you ask c
by XCSme 3mo ago
It is on top for many benchmarks, only not the coding/agentic ones.
Still one of the most intelligent models overall, most likely to get any question you ask correctly (without tools).
- Zababa 3mo agoNot in my experience, it tends to pick up subtle orientations given in a question (like "which is better, A and B?" and in the context you add you list a few things for A and B) and will absolutely run with them even if they're not true. Has been an issue with Gemini models since at least 3.0. Maybe that makes them great roleplaying models, but for factual information they just run with the slightest hint in one direction or another and never really push back objectively.