6 ms·
GPT-5.2 and GPT-5.2-Codex are now 40% faster
- simianwords 8mo agoIt’s interesting that they kept the price the same while doing inference on Cerebras is much more expensive.
- chillee 8mo agothis is almost certainly not being done on cerebras
- diwank 8mo agoI dont think this is Cerebras. Running on cerebras would change model behavior a bit and it could potentially get a ~10x speedup and it'd be more expensive. So most likely this is them writing new more optimized kernels for Blackwell series maybe?
- simianwords 8mo agoFair point but it remains to answer - why isn’t this speed up available in ChatGPT and only in the api?
- angoragoats 8mo ago[flagged]
- deleted 8mo ago[deleted]
- prodigycorp 8mo agoThis is great. In the past month, OpenAI has released for codex users: - subagents support - a better multi agent interface (codex app) - 40% faster inference No joke, with the first two my productivity is already up like 3x. I am so stoked to try this out.
- brianwawok 8mo agoTry Claude and you can get x^2 performance. OpenAI is sweating
- klipklop 8mo ago5.2-codex is pretty solid and you get dramatically higher usage rates with cheap plans. I would assume API use is much cheaper as well.
- jerkstate 8mo agopeople are sleeping on openai right now but codex 5.2 xhigh is at least as good as opus and you get a TON more usage out of the OpenAI $20/mo plan than Claude's $20/mo plan. I'm always hitting the 5 hour quota with Opus but never have with Codex. Codex tool itself is not quite as good but close.
- indemnity 8mo agoIs there a plan like the $100 Claude Max? $200 for ChatGPT Pro is a little bit too much for me. Whereas Claude Max 5x is enough that I don’t really run out with my usage patterns.
- jerkstate 8mo agoIf $20/mo Claude is not enough for you but 5x Claude at $100/mo is, the $20 chatgpt plus subscription might give you enough codex for your usage
- viraptor 8mo agoMay be a bit different depending on what kind of work you're doing, but for me 5.2-codex finally reached higher level than opus.
- akmarinov 8mo agoIf i could use GPT-5.2 with Claude Code - yeah. Otherwise slOpus requires too much steering to get things done. GPT-5.2 just works
- OutOfHere 8mo agoOpenAI in my estimation has the habit of dropping a model's quality after its introduction. I definitely recall the web ChatGPT 5.2 being a lot better when it was introduced. A week or two later, its quality suddenly dropped. The initial high looked to be to throw off journalists and benchmarks. As such, nothing that OpenAI says in terms of model speed can be trusted. All they have to do is lower the reasoning effort on average, and boom, it becomes 40% faster. I hope I am wrong, because if I am right, it's a con game. Starting off the ChatGPT Plus web users with the Pro model, then later swapping it for the Standard model -- would meet the claims of model behavior consistency, while still qualifying as shenanigans.
- bethekidyouwant 8mo agoI mean you can just run the benchmark again
- OutOfHere 8mo agoHow are you going to benchmark the web ChatGPT Plus, which is where a reduction was suspected?
- tedsanders 8mo agoIt's good to be skeptical, but I'm happy to share that we don't pull shenanigans like this. We actually take quite a bit of care to report evals fairly, keep API model behavior constant, and track down reports of degraded performance in case we've accidentally introduced bugs. If we were degrading model behavior, it would be pretty easy to catch us with evals against our API. In this particular case, I'm happy to report that the speedup is time per token, so it's not a gimmick from outputting fewer tokens at lower reasoning effort. Model weights and quality remain the same.
- wahnfrieden 8mo agoYou're confirming you don't alter "juice" levels..?
- riku_iki 8mo agotons of posts on reddit that they also significantly dropped quality
- samusiam 8mo agoThere are always people on reddit saying such-and-such model quality significantly dropped. Every single day there's a post like this in one of the Claude sub-reddits. It's virtually never substantiated with reliable evidence.
- thadk 8mo agoIt was probably from the other day when roon realized that normal people have it slower than staff. Then from that they realized they could just run API calls more like staff, fast, not at capacity. Then they leave the billion other people's calls at remaining capacity. https://thezvi.substack.com/i/185423735/choose-your-fighter https://thezvi.substack.com/i/185423735/choose-your-fighter > Ohqay: Do you get faster speeds on your work account? > roon: yea it’s super fast bc im sure we’re not running internal deployment at full load
- thebigspacefuck 8mo agoSpeed was always my main complaint, these models always felt really good but too slow. I’ll have to give them a try again.
- tmaly 8mo agoOver the weekend I was running the same prompt across GPT-5.2, Gemini 3, and Grok. Both Gemini 3 and Grok on thinking mode finished within 2 minutes. GPT-5.2 was just spinning its wheels for like 6 minutes.
- logicallee 8mo agoany ideas how they could get the speedup?