3 ms·
Is there a good benchmark leaderboard between coding agents?
by jadbox 4mo ago
Is there a good benchmark leaderboard between coding agents?
- Imustaskforhelp 4mo agohttps://artificialanalysis.ai/agents/coding-agents?coding-agents-performance-chart=index https://artificialanalysis.ai/agents/coding-agents?coding-ag... This seems to be a benchmark but sadly between just primarily claude-code, codex,cursor and (gemini-cli?)
- impulser_ 4mo agoHarnesses aren't really going to change much of the performance on models like Opus, and GPT. You literally can just give the model a bash tool and it will do just fine in fact it will most likely do better than majority of harnesses due to how well models are at bash. The model do all the lifting. It really doesn't matter which harness you use.