3 ms·
What's interesting, is that Sonnet 5 is actually worse[0] than 4.6 without reasoning. It makes some sense, as models are trained more and more with reasoning,
by XCSme 3mo ago
What's interesting, is that Sonnet 5 is actually worse[0] than 4.6 without reasoning.
It makes some sense, as models are trained more and more with reasoning, than without.
[0]: https://aibenchy.com/compare/anthropic-claude-sonnet-4-6-none/anthropic-claude-sonnet-5-none/ https://aibenchy.com/compare/anthropic-claude-sonnet-4-6-non...