5 ms·
You should try a better harness. Try pi, or ohmypi if you want a good OOB experience
by rattlesnakedave 2mo ago
You should try a better harness. Try pi, or ohmypi if you want a good OOB experience
- gigatexal 2mo agoI’m in the Claude code harness for everything boat too. What are the alternatives?
- KronisLV 2mo agoWhat the person above is suggesting: * https://pi.dev/ https://pi.dev/ * https://omp.sh/ https://omp.sh/ (no personal opinions of either, links might be useful) I think that OpenCode is nice, their CLI version is enjoyable and their desktop/web version is okay: * https://opencode.ai/ https://opencode.ai/ I also quite like driving OpenCode through something like Kepler / Paseo and tools like that (with those I can still use my Anthropic Condition by Claude Code being treated similarly - as something that gets tasks dispatched to it, while the GUI I see is Kepler / Paseo). On the desktop side, ZCode was surprisingly usable for something that came out of nowhere (I wasn't aware of it at all before trying out the GLM Coding Plan): https://zcode.z.ai/en https://zcode.z.ai/en
- vadansky 2mo agoLast time I tried some of these, none of them had the "manual mode" that CC has, where it shows you change by change as diffs and you can edit them before accepting and moving on to the next change. I like that because if it's going off pattern I can spot it early on and guide it correctly, instead of having to review the whole completed diff at the end when it's too late. I should spend the weekend checking them out again to see if they added that but I assume with everyone going full agent mode they probably didn't.
- eli 2mo agoThe philosophy with Pi is it is minimal (but functional) out of the box and easily extensible. I'm not familiar with that feature but I would not at all be surprised if someone already coded a Pi extension that does it.
- 0xbadcafebee 2mo agoBoth Pi and OpenCode let you customize them. You tell the AI you want "something like claude code manual mode", and they'll modify your configs to do the same thing, or build an extension for you (however, it's much faster to use Plan Mode to build a plan of what it will do, and then execute the plan in Build Mode. you can also have the AI make a script that will be executed deterministically)
- badcafe23423435 2mo agoNo one in their right mind would install software using `curl | bash`
- schaefer 2mo agoI feel the same, which is why I only use these tools in a docker container. Because life is compromise.
- redeeman 2mo agoits a good way to check if people are insane though. would be a cool tactic for new hire evaluation, monitor them setting up dev environment. do the curl | bash, and its instafail
- chrisweekly 2mo agoIn countless corporate environments (including in highly regulated industries), far from being a firable offense, piping curl to bash is often a prescribed step in setting up the standard dev env. The cognitive dissonance is soul crushing. Maybe they're testing for one's ability to tolerate it.
- cjbprime 2mo agoWhy not? The web uses TLS, how's it different security-wise compared to a package download?
- bityard 2mo agoIt's less about malicious intent and more about predictability. When I install software on my computer with apt, I trust that all the files will go to the right place and install scripts are going to do sane things relative to the rest of the system. And I can just uninstall the whole thing with one command later if I so choose. If I curlpipe a script, I get none of those guarantees. I have seen curlpipes that put files in weird places, guess the wrong OS, and mess with config files that I didn't want them to touch. When they break or I want to uninstall, I have to sit down and understand a (possibly minified) script to clean things up manually. Yes containers are a half solution to this, no I don't want to use containers 100% of the time.
- gigatexal 2mo agoThe out of box experience of omp.sh is wow imo so much nicer than Claude. Claude spends too much time being nice and gassing me Up and omp just gets to work. It’s idk smarter like a far better system prompt and all around loop. Gonna try to find a way to use this at work.
- gigatexal 2mo agoBy the way is anyone running any security / audit / exfil tests on these open harnesses? Like a VPN id kinda like to know? I’m happy to help fund a crowd source campaign for it.
- nextaccountic 2mo agohttps://github.com/tontinton/maki https://github.com/tontinton/maki is tackling the right issues IMO. not sure how they compare with the rest
- aqme28 2mo agoArtificialAnalysis puts out benchmarks for harnesses now as well, and OpenCode seems to be winning it. https://artificialanalysis.ai/agents/coding-agents#coding-agents-harness-comparison-chart-tabs https://artificialanalysis.ai/agents/coding-agents#coding-ag... I only found this yesterday, and it inspired me to start testing out OpenCode.
- dmix 2mo agoI'm not surprised to also see Cursor above Claude code, their harness is very good.
- leobuskin 2mo agoIn what scenarios?
- dmix 2mo agoIt indexes the code efficiently, seems to find stuff quicker, it has a very nice UI (much better than Claude Codes IMO), it has a nice sub-agent UX which I find triggers more reliably, diffs render nicely. Otherwise it just seems to work in a purely vibes sense. That said Claude Code is perfectly fine. I just prefer the integrated experience of using Cursors since I already use VSCode, but I still mostly use Claude Code because of their Max/Fable plan.
- bensyverson 2mo agoCan OpenCode dispatch background subagents yet? I tried it a week ago and saw nothing. This is 99% of my workflow at this point.
- aqme28 2mo agoHaving just "asked" my opencode instance-- Yes, but it's behind an experimental flag and not necessarily feature-complete.
- jauntywundrkind 2mo ago
- leobuskin 2mo agoI’ve tried a bunch of them, and I seriously do not understand these recommendations. It was a rough road and a steep hill, but right now CC is absolutely the best harness on the market, as for me, whatever top tier model is under the hood (mostly, some of them, like DeepSeek, don’t fit CC at all).
- infecto 2mo agoInversely I don’t understand the praise for CC. These days it feels like bloatware. It absolutely can get the work done but when I measure on token and time use it ends up being a multiple of pi like harnesses. CC works but for me it felt like increasingly they have zero incentive to make it a great experience. You hear folks like Boris talk about spinning up thousands of agents over night and agents chatting back and forth in GitHub issues and while I think it’s great from figuring out what the future looks like I don’t think it represents the reality of ROI today. So the folks building the tool are so disconnected I am simply not sure it’s a great experience anymore.
- barbazoo 2mo agoSo is the quantitative difference in token use the only difference or do you think there's also a different qualitat? I'm on CC only and immensely happy. Very productive both at work and privately and at work I average around $250 a month which probably means nothing but it's little compared to my salary. Is that the main concern though, cost?
- disgruntledphd2 2mo agoFor me, at least it's that the newer Claude models seem optimised for one-shotting things, which is not what I want. As the amount of code per turn increases, I have a harder job keeping up and ensuring that it's doing what I want. That being said, I had to nope out of a similar thing from GPT 5.6 today, so it appears to be a US frontier lab issue. Claude is particularly bad though, as it produces far too much code even when I tell it not to, unlike GPT (and Kimi) which at least listen to me a little better. More generally, I want a usable human review experience, and Claude code doesn't deliver that for me.
- gigatexal 2mo agothank you all! got something to tinker with this weekend i like to challenge my assumptions and try new tools
- infecto 2mo agoJust as a +1 anecdote. I enjoy using pi a lot. I used to h think the harness matters a lot but with the current iteration of models I am starting to sway that while it matters it’s less and less important and that CC is bloated. I did some quick tests when I switched and a task that would take $5 in tokens would be completed in $0.50 in pi. Very anecdotal and I don’t have a test framework setup to make this very official but increasingly felt like CC was spinning its wheels on the easiest of tasks.
- xabd 2mo agoThe token cost difference is pretty interesting. I wonder how much of that is the harness itself versus how aggressively each one loops, plans, and calls tools. A proper apples-to-apples test would be really useful here.
- gigatexal 2mo agoare ohmypi and pi related? that's a very compelling use case, thank you
- agentdev001 2mo agoYes, ohmypi is an opinionated set of features on the base pi harness.
- yumosx 2mo ago[dead]
- skybrian 2mo agoIf you like running everything in a VM and using a web browser as your UI, Shelley is very good: https://github.com/boldsoftware/shelley https://github.com/boldsoftware/shelley It works nicely in the browsers on my tablet and phone, too. On exe.dev you can ask it to customize itself, and it will automatically rebase your customizations when upgrading to a new release.
- rob 2mo agoT3 Code has been amazing. Completely free. Really impressed with the desktop app and the mobile app experience and the way it works seamlessly has me actually accomplishing tons of stuff while I'm out on mobile that I would otherwise have to wait to come home for. First time in a while I'm actually excited to use a desktop UI instead of the terminal. Blows away the official Claude Code mobile app. I can switch between my Claude and Codex monthly subscriptions in it as well. There's a TestFlight beta SwiftUI mobile version that's so much nicer than the one in the App Store. I'm running the nightly version of the desktop app. And this is coming from someone that's not particularly a big fan of Theo. T3 Code should get more recognition; people aren't just aware of it yet.
- schmuhblaster 2mo agoI think that writing your own harness is a rite of passage now, just like writing your own search engine or database, rolling your own crypto… Anyways, please try mine! https://github.com/deepclause/deepclause-sdk https://github.com/deepclause/deepclause-sdk
- DenisM 2mo agoeventually we will get to “just use postrges” stage
- jpadkins 2mo agoPiggybacking on this thread to ask my question: What are alternatives that are multiplayer (team oriented) by default? For example, I want my team to see all my sessions easily, vise versa. another way of stating: all the agents are running in a container that that any member of the team can view and interact with.
- rush86999 2mo agoMine is a WIP for automation but has similar concepts to what you're looking for: https://github.com/rush86999/atom https://github.com/rush86999/atom
- vorticalbox 2mo agoIt’s not just bloat at this point. I run oMLX and run models locally. using Claude code on the first message dumps 40k of tokens that my laptop takes 5 mins to compute. I’ve stopped using it completely now.
- dominotw 2mo agowhat is this comment based on ? vibes?
- makerdiety 2mo agobecause not everything is a shilling advertisement?
- infecto 2mo agoVibes like your low quality comment? What’s the counter argument? pi and ohmypi are pretty fantastic. Of course like all developer tools it depends how you do your work but I am not sure what you are trying to achieve in your comment.
- rpdillon 2mo agoBased on the fact that Claude Code is only optimized for Anthropic models, whereas Pi and Omp are optimized for a wide variety of models, including open weights.
- dominotw 2mo agothey are not really optimized for 'wide variety of models' . what optimization did pi do for glm 5.3?
- yogthos 2mo agoI'm gonna shamelessly plug my own here :) https://dirge-code.github.io/ https://dirge-code.github.io/
- Shorel 2mo agoI like it. I think creating your own agent is the Hello World of agentic coding. Instead of Rust, I used D for mine.
- anentropic 2mo agoThat actually looks nice
- johng 2mo agoI get hung here on Debian 13 after installing rustup and doing rustup install stable. Building [=======================> ] 610/611: dirge(bin) Just hangs there :(
- yogthos 2mo agomight just be slow, it can take over 5 min to compile
- weego 2mo agowhat is a harness? The comments below are mixing IDE/ADE but other suggestions are purely terminal things and I don't get what their value is over just a terminal. Is a harness like a loop where it's just a vague thing that everyone nods about but everyone is nodding at something different?
- jewel 2mo agoMy understanding is that the harness is the set of function calls (or tool calls) that let the LLM interact with your codebase. It's independent of the IDE or CLI. The tool calls will be, among other things, something like ReadFile, RipGrep, PatchFile, Shell. When people talk about the value of different harnesses, they're also implicitly talking about the quality of the system prompt. The same exact model, when given a different set of tools and a different system prompt, can behave differently.
- kristjansson 2mo agoharnesss == thing that calls LLM API, acts on response, and maybe does that again.
- computerex 2mo agoThe harness is the agent. LLM's can be asked to output things in JSON for example. The LLM then literally asks for things like "execute this cmd" or search/replace this string. The LLM outputs text, but in a deterministic format that can be parsed. The harness calls the LLM, exposes tools, executes tools the LLM asks for, gates tool use based on security controls. It's the runtime that the agent uses to do work.
- stogot 2mo agoAre the LLM and agent the same thing? Why different nouns ?
- computerex 2mo agoThe LLM is the core model, but the harness has the prompts/tool definitions, guidance/recovery/correction code. The harness itself is the agent, because same model may perform vastly differently on different harnesses. Agent is the system working as a whole, harness+llm.
- PeterStuer 2mo agoYou sound like I could afford that.
- vinhnx 2mo agoI have been building and maintaining a coding harness, with help from the community https://github.com/vinhnx/VTCode https://github.com/vinhnx/VTCode. Hope you'll check it out.