5 ms·
I’ve updated my article (replaced GPT-5 with ChatGPT-5 in this section): When you ask the latest model, ChatGPT-5 to multiply two large numbers, it doesn't cal
by QueensGambit 1y ago
I’ve updated my article (replaced GPT-5 with ChatGPT-5 in this section):
When you ask the latest model, ChatGPT-5 to multiply two large numbers, it doesn't calculate. It generates Python code, executes it in a sandbox, and returns the result. Unlike ChatGPT-3, which at least attempted arithmetic internally (and often failed), ChatGPT-5 delegates computation to external tools. [1]
And added this note:
[1] There are 2 ways to multiply numbers in GPT-5:
- Python mode, which uses python sandbox as mentioned above
- No tool mode, which uses internal reasoning
Python mode is approximately 2x more accurate than no tool mode in FrontierMath (26.3% vs 13.5% accuracy on expert level math). Python mode is also 4x to 10x more cost effective than no tool mode.
The GPT-5 API uses no-tool mode by default (tools must be explicitly enabled in API calls), while ChatGPT UI likely uses Python mode by default since Advanced Data Analysis is enabled by default for all Plus, Team, and Enterprise subscribers. This creates a significant cost optimization for OpenAI in the consumer product, while API users bear the full cost of inefficient reasoning unless they manually configure tool use.
---
Thanks again for flagging the inaccuracy, Simon! If you think any part of this update still misrepresents the model behavior, I’d love your input.