3 ms·
It definitely feels like subscription limits have dropped since GPT-5.6 came out. But I couldn't really find any evidence for it, going through my history. All
by vincent_s 2mo ago
It definitely feels like subscription limits have dropped since GPT-5.6 came out. But I couldn't really find any evidence for it, going through my history. All I could come up with is that 5.6-sol uses about twice the tokens compared to 5.5 (both on xhigh) [0].
One thing I just did was to stop four long-running Codex sessions that ran on 5.6-sol and switched to 5.5 and asked it "Please check if you're really going against the actual goal or have you drifted away from that?" and all four replied something like "Yes: I had started to drift"
[0] https://www.vincentschmalbach.com/gpt-5-6-sol-xhigh-uses-twice-tokens-gpt-5-5/ https://www.vincentschmalbach.com/gpt-5-6-sol-xhigh-uses-twi...
- vincent_s 2mo agoI have seen that kind of drift in smaller models (e.g. DeepSeek V4 Flash) when setting the thinking too high. So less thinking would lead to better results. But that's not something I'd expect from a SOTA model. Higher thinking effort should lead to same or better results.
- esperent 2mo agoI always set thinking to high for the main session or subagents doing audits, medium for everything else. I think xhigh or max has a place if I'm doing something genuinely complex. But in general use it's slow and I have the impression the highest often gives worse results - over-thought, over-engineered. High or medium might give better results for less complex tasks.