4 ms·
> “a new knowledge cutoff of August 2025” This (and the price increase) points to a new pretrained model under-the-hood. GPT-5.1, in contrast, was allegedly u
by jumploops 10mo ago
> “a new knowledge cutoff of August 2025”
This (and the price increase) points to a new pretrained model under-the-hood.
GPT-5.1, in contrast, was allegedly using the same pretraining as GPT-4o.
- 98Windows 10mo agoor maybe 5.1 was an older checkpoint and has more quantization
- deleted 10mo ago[deleted]
- FergusArgyll 10mo agoA new pretrain would definitely get more than a .1 version bump & would get a whole lot more hype I'd think. They're expensive to do!
- femiagbabiaka 10mo agoNot if they didn't feel that it delivered customer value no? It's about under promising and over delivering, in every instance
- redwood 10mo agoNot if it underwhelms
- hannesfur 10mo agoMaybe they felt the increase in capability is not worth of a bigger version bump. Additionally pre-training isn't as important as it used to be. Most of the advances we see now probably come from the RL stage.
- caconym_ 10mo agoReleasing anything as "GPT-6" which doesn't provide a generational leap in performance would be a PR nightmare for them, especially after the underwhelming release of GPT-5. I don't think it really matters what's under the hood. People expect model "versions" to be indexed on performance.
- ACCount37 10mo agoNot necessarily. GPT-4.5 was a new pretrain on top of a sizeable raw model scale bump, and only got 0.5 - because the gains from reasoning training in o-series overshadowed GPT-4.5's natural advantage over GPT-4. OpenAI might have learned not to overhype. They already shipped GPT-5 - which was only an incremental upgrade over o3, and was received poorly, with this being a part of the reason why.
- diego_sandoval 10mo agoI jumped straight from 4o (free user) into GPT-5 (paid user). It was a generational leap if there ever has been one. Much bigger than 3.5 to 4.
- boc 10mo agoMaybe the rumors about failed training runs weren't wrong...
- jumploops 10mo agoIt’s possible they’re using some new architecture to get more up-to-date data, but I think that’d be even more of a headline. My hunch is that this is the same 5.1 post-training on a new pretrained base. Likely rushed out the door faster than they initially expected/planned.
- OrangeMusic 10mo agoYeah because OpenAI has been great at naming their models so far? ;)
- MagicMoonlight 10mo agoNo, they just feed in another round of slop to the same model.
- redox99 10mo agoI think it's more likely to be the old base model checkpoint further trained on additional data.