3 ms·
I wonder why they didn't compare with GPT-5.6-sol, only Terra?
by ipsum2 2mo ago
I wonder why they didn't compare with GPT-5.6-sol, only Terra?
- woadwarrior01 2mo agoHaven't you seen the kernel optimization case study at the bottom of the page? They compare against GPT-5.6 Sol and their model is worse.
- wmf 2mo agoClearly they're positioning it as a mid model.
- minimaxir 2mo agoWhich is in itself a bit weird as mid models nowadays are a golden mean fallacy. Terra is much less popular than both Luna (cost-sensitive) and Sol (performance-sensitive). Claude Sonnet is a weird exception to the mid models because Anthropic doesn't do much with Haiku and Opus is too big.
- ukblewis 2mo agoI don’t know where you get your statistics, but I love Terra and use it all of the time. It is the default fastest model in ChatGPT/Codex today. I saw today a notice saying that the model had hit capacity briefly
- deaux 2mo agoIt's a little of the opposite to what you're saying. By performance you seem to mean only "intelligence" i.e. pass rate. But there's a 3rd factor, time to task completion. So it's three-dimensional rather than two. And on that spectrum there are areas where Terra is optimal. Sonnet on the other hand never is, it's far from the best pick anywhere on the spectrum.
- redox99 2mo agoBut why include Opus then?
- logicchains 2mo agoPresumably because it's worse than Sol, same reason they compared it to Opus 5 not Fable.
- Handy-Man 2mo agoTheir bigger model is not ready - watermelon code name was still being prepared for release as of a month ago