4 ms·
> only Claude Code has been able to really solve; Sonnet 4.5 in there consistently performs better than Sonnet 4.5 anywhere else. I think part of it is this[0]
by mritchie712 11mo ago
> only Claude Code has been able to really solve; Sonnet 4.5 in there consistently performs better than Sonnet 4.5 anywhere else.
I think part of it is this[0] and I expect it will become more of a problem.
Claude models have built-in tools (e.g. `str_replace_editor`) which they've been trained to use. These tools don't exist in Cursor, but claude really wants to use them.
0 - https://x.com/thisritchie/status/1944038132665454841?s=20 https://x.com/thisritchie/status/1944038132665454841?s=20
- HugoDias 11mo agoTIL! I'll finally give Claude Code a try. I've been using Cursor since it launched and never tried anything else. The terminal UI didn't appeal to me, but knowing it has better performance, I'll check it out. Cursor has been a terrible experience lately, regardless of the model. Sometimes for the same task, I need to try with Sonnet 4.5, ChatGPT 5.1 Codex, Gemini Pro 3... and most times, none managed to do the work, and I end up doing it myself. At least I’m coding more again, lol
- firloop 11mo agoYou can install the Claude Code VS Code extension in Cursor and you get a similar AI side pane as the main Cursor composer.
- adastra22 11mo agoThat’s just Claude Code then. Why use cursor?
- dcre 11mo agoPeople like the tab completion model in Cursor.
- BoorishBears 11mo agoAnd they killed Supermaven. I've actually been working on porting the tab completion from Cursor to Zed, and eventually IntelliJ, for fun It shows exactly why their tab completion is so much better than everyone else's though: it's practically a state machine that's getting updated with diffs on every change and every file you're working with. (also a bit of a privacy nightmare if you care about that though)
- fragmede 11mo agoit's not about the terminal, but about decoupling yourself from looking at the code. The Claude app lets you interact with a github repo from your phone.
- verdverm 11mo agoThis is not the way these agents are not up to the task of writing production level code at any meaningful scale looking forward to high paying gigs to go in and clean up after people take them too far and the hype cycle fades --- I recommend the opposite, work on custom agents so you have a better understanding of how these things work and fail. Get deep in the code to understand how context and values flow and get presented within the system.
- fragmede 11mo ago> these agents are not up to the task of writing production level code at any meaningful scale I think the new one is. I could be the fool and be proven wrong though.
- verdverm 11mo agoIt's marginally better, no where close to game changing, which I agree will require moving beyond transformers to something we don't know yet
- alwillis 11mo ago> these agents are not up to the task of writing production level code at any meaningful scale This is obviously not true, starting with the AI companies themselves. It's like the old saying "half of all advertising doesn't work; we just don't which half that is." Some organizations are having great results, while some are not. From the multiple dev podcasts I've listened to by AI skeptics have had a lightbulb moment where they get AI is where everything is headed.
- verdverm 11mo ago
- idonotknowwhy 11mo agoGlad you mentioned "Cursor has been a terrible experience lately", as I was planning to finally give it a try. I'd heard it has the best auto-complete, which I don't get use VSCode with Claude Code in the terminal.
- alphabettsy 11mo agoYou should still give it a try. Can’t speak for their experience, but doesn’t ring true for me.
- baq 11mo ago+1, it had a bad period when they were hyperscaling up, but IME they've found their pace (very) recently - I almost ditched cursor in the summer, but am a quite happy user now.
- PrayagS 10mo agoI haven’t used Cursor since I use Neovim and it’s hard to move out. The auto-complete suggestions from FIM models (either open source or even something Gemini Flash) punch far above their weight. That combined with CC/Codex has been a good setup for me.
- fabbbbb 11mo agoI get the same impression. Even GPT 5.1 Codex is just sooo slow in Cursor. Claude Code with Sonnet is still the benchmkar. Fast and good.
- Huppie 11mo agoI was evaluating codex vs claude code the past month and GPT 5.1 codex being slow is just the default experience I had with it. The answers were mostly on par (though different in style which took some getting used to) but the speed was a big downer for me. I really wanted to give it an honest try but went back to Claude Code within two weeks.
- bgrainger 11mo agoThis feels like a dumb question, but why doesn't Cursor implement that tool? I built my own simple coding agent six months ago, and I implemented str_replace_based_edit_tool (https://platform.claude.com/docs/en/agents-and-tools/tool-use/text-editor-tool#claude-4 https://platform.claude.com/docs/en/agents-and-tools/tool-us...) for Claude to use; it wasn't hard to do.
- svnt 11mo agoMaybe this is a flippant response, but I guess they are more of a UI company and want to avoid competing with the frontier model companies? They also can’t get at the models directly enough, so anything they layer in would seem guaranteed to underperform and/or consume context instead of potentially relieving that pressure. Any LLM-adjacent infrastructure they invest in risks being obviated before they can get users to notice/use it.
- fabbbbb 11mo agoThey did release the Composer model and people praise the speed of it.
- paradite 11mo agoMaybe they want to have their own protocol and standard for file editing for training and fine-tuning their own models, instead of relying on Anthropic standard. Or it could be a sunk cost associated with Cursor already having terabytes of training data with old edit tool.
- ayewo 11mo agoIs the code to your agent and its implementation of "str_replace_based_edit_tool" public anywhere? If not, can you share it in a Gist?