5 ms·
Thanks for the video. His fix for "the dumb zone" is the RPI Framework: ● RESEARCH. Don't code yet. Let the agent scan the files first. Docs lie. Code doesn't
by kaizenb 7mo ago
Thanks for the video.
His fix for "the dumb zone" is the RPI Framework:
● RESEARCH. Don't code yet. Let the agent scan the files first. Docs lie. Code doesn't.
● PLAN. The agent writes a detailed step-by-step plan. You review and approve the plan, not just the output. Dex calls this avoiding "outsourcing your thinking." The plan is where intent gets compressed before execution starts.
● IMPLEMENT. Execute in a fresh context window.
The meta-principle he calls Frequent Intentional Compaction: don't let the chat run long. Ask the agent to summarize state, open a new chat with that summary, keep the model in the smart zone.
- girvo 7mo agoThat's fascinating: that is identical to the workflow I've landed on myself.
- hedora 7mo agoIt's also identical to what Claude Code does if you put it in plan mode (bound to <tab> key), at least in my experience.
- girvo 7mo agoMy annoyance with plan mode is where it sticks the .md file, kind of hides it away which makes it annoying to clear context and start up a new phase from the PLAN file. But that might just be a skill issue on my end
- hedora 7mo agoEven worse, it just randomly blows away the plan file without asking for permission. No idea what they were thinking when they designed this feature. The plan file names are randomly generated, so it could just keep making new ones forever for free (it would take a LONG time for the disk space to matter), but instead, for long plans, I have to back the plan file up if it gets stuck. Otherwise, I say "You should take approach X to fix this bug", it drops into plan mode, says "This is a completely unrelated plan", then deletes all record of what it was doing before getting stuck.
- girvo 7mo agoIt’s not just me then! Hah good to know. It’s why I’ve started ignoring plan modes in most agent harnesses, and managing it myself through prompting and keeping it in the code base (but not committed)
- toddmerrill 7mo agoMy experience also. The claude code document feature is a real missed opportunity. As you can see in this discussion, we all have to do it manually if we want it to work.
- kaizenb 7mo agoAfter creating the plan in Plan mode (+Thinking) I ask Claude to move the plan .md file to /docs/plans folder inside the repo. Open a new chat with Opus, thinking mode is off. Because no need when we have detailed plan. Now the plan file is always reachable, so when the context limit is narrowing, mostly around 50%, I ask Claude to update the plan with the progress, and move to a new chat @pointing the plan file and it continue executing without any issue.
- insane_dreamer 7mo agobetter to instruct it to write a plan .md file that is appropriately named so that it can be easily referenced/updated in multiple sessions. I've found that effective.
- cruffle_duffle 7mo agoDunno if you know this but the plan in plan mode is a markdown file! Ask it for the file and it will give it to you.
- insane_dreamer 7mo agoyes, but if you start a fresh session to continue working on your project, it's a lot easier if you already know which PLAN file you need for your project. Plus you can commit it.
- cortesoft 7mo agoIt’s the style spec-kit uses: https://github.com/github/spec-kit https://github.com/github/spec-kit Working on my first project with it… so far so good.
- iamacyborg 7mo ago> RESEARCH. Don't code yet. Let the agent scan the files first. Docs lie. Code doesn't. I find myself often running validity checks between docs and code and addressing gaps as they appear to ensure the docs don’t actually lie.
- silverlake 7mo agoI have Codex and Gemini critique the plan and generate their plans. Then I have Claude review the other plans and add their good ideas. It frequently improves the plan. I then do my careful review.
- ArtRichards 7mo agoThis is exactly how I've found leads to most consistent high quality results as well. I don't use gemini yet (except for deep research, where it pulls WAY ahead of either of the other 'grounding' methods) But Codex to plan big features and Claude to review the feature plan (often finds overlooked discrepancies) then review the milestones and plan implementation of them in planning mode, then clear context and code. Works great.
- Huppie 7mo agoMore recently I've been doing the implement phase without resetting the whole context when context is still < 60% full and must say I find it to be a better workflow in many cases (depends a bit on the size of the plan I suppose.) It's faster because it has already read most relevant files, still has the caveats / discussion from the research phase in its context window, etc. With the context clear the plan may be good / thorough but I've had one too many times that key choices from the research phase didn't persist because halfway through implementation Opus runs into an issue and says "You know what? I know a simpler solution." and continues down a path I explicitly voted down.
- greenchair 7mo agoHow is that Plan strategy not "outsourcing your thinking" because that's exactly what it sounds like. AI does the heavy lifting and you are the editor.
- brookst 7mo agoIs a VP of engineering “outsourcing their thinking” by having an org that can plan and write software?
- Filligree 7mo agoYes.
- Eldt 7mo agoDelegation is generally all about outsourcing, so hard agree
- brookst 7mo agoInteresting take. Does that mean SWE's are outsourcing their thinking by relying on management to run the company, designers to do UX, support folks to handle customers? Or is thinking about source code line by line the only valid form of thinking in the world?
- qualifck 7mo agoI mean yes? That's like, the whole idea behind having a team. The art guy doesn't want to think about code, the coder doesn't want to think about finances, the accountant doesn't want to worry about customer support. It would be kind of a structural failure if you weren't outsourcing at least some of your thinking.
- brookst 7mo agoI’m with you, perhaps I just misread some kind of condescension into the “outsourcing your thinking” comment. We all have limited context windows, the world’s always worked that way, just seemed odd to (mis)read someone saying there’s something wrong with focusing on when you add the greatest value and trusting others to do the same.
- dahart 7mo agoAdd a REFLECT phase after IMPLEMENT. I’m finding it’s extremely useful to ask agents for implementation notes and for code reviews. These are different things, and when I ask for implementation notes I get very different output than the implementation summary it spits out automatically. I ask the agent to surface all design choices it had to make that we didn’t explicitly discuss in the plan, and then check in the plan + impl notes in order to help preload context for the next thing. My team has been adopting a separation of plan & implement organically, we just noticed we got better output that way, plus Claude now suggests in plan mode to clear context first before implementing. We are starting to do team reviews on the plan before the implement phase. It’s often helpful to get more eyeballs on the plan and improve it.