6 ms·
I don’t know what the people who say Claude 3 is better than GPT-4 are using it for. It’s been consistently worse for everything I’ve thrown at it. Debugging a
by thorum 3y ago
I don’t know what the people who say Claude 3 is better than GPT-4 are using it for. It’s been consistently worse for everything I’ve thrown at it.
Debugging a Python function this morning. Claude 3 Opus failed completely. GPT-4 found the bug, as well as two others I hadn’t even been looking for.
- simonw 3y agoI've had the opposite experience: coding prompts that GPT-4 makes mistakes on Claude 3 Opus gets right the first time. As always, your results will vary based on your personal prompting style. My style apparently works great with Opus. Here's one example: GPT-4 gave me code that was missing some async/await keywords: https://chat.openai.com/share/117fb1ad-6361-41e2-be59-110f3262594d https://chat.openai.com/share/117fb1ad-6361-41e2-be59-110f32... Claude 3 Opus with the same prompt got it right the first time: https://gist.github.com/simonw/2002e2b56a97053bd9302a34e0b83074 https://gist.github.com/simonw/2002e2b56a97053bd9302a34e0b83...
- brianjking 3y agoYeah, Opus has entirely taken over any code specific use for me over ChatGPT 4 or OpenAI GPT-4 API. Once Opus has the ability to run a code interpreter, it'll really be an exciting time.