5 ms·
The charts are also extremely difficult to parse. They seem auto-generated. Dataset coloring is atrocious. Regarding your main point, yes, I agree. My impressi
by enraged_camel 3mo ago
The charts are also extremely difficult to parse. They seem auto-generated. Dataset coloring is atrocious.
Regarding your main point, yes, I agree. My impression (as someone who uses both Codex and Claude Code daily) is that OpenAI does a fair amount of benchmaxxing.