3 ms·
> As a counter I’ve had OpenAI Codex and Claude Code both catch logic cases I’d missed in both tests and codes That has other explanations than that it reasone
by AstroBen 9mo ago
> As a counter I’ve had OpenAI Codex and Claude Code both catch logic cases I’d missed in both tests and codes
That has other explanations than that it reasoned its way to the correct answers. Maybe it had very similar code in its training data
This specific example was with Codex. I didn't mention it because I didn't want it to sound like I think codex is worse than claude code
I do realize my prompt wasn't optimal to get the best out of AI here, and I improved it on the second pass, mainly to give it more explicit instruction on what to do
My point though is that I feel these situations are heavily indicative of it not having true reasoning and understanding of the goals presented to it
Why can it sometimes catch the logic cases you miss, such as in your case, and then utterly fail at something that a simple understanding of the problem and thinking it through would solve? The only explanation I have is that it's not using actual reasoning to solve the problems