3 ms·
> I've never had an issue with Codex or Claude reading massive files Reading files isn't a problem they want to solve. The idea seems to be using a cheaper mod
by jampa 1mo ago
> I've never had an issue with Codex or Claude reading massive files
Reading files isn't a problem they want to solve. The idea seems to be using a cheaper model to "scout" for the intended code, instead of an expensive one that reads all the things (and spends more tokens / thinks about them).
I think this might be useful because Opus 5 especially tends to over-read. So this looks like an "LLM Bloom filter", telling "hey this is the code you might want to read".
- hankbond 1mo ago> "LLM Bloom filter" very good way to put it.
- lxgr 29d agoIt's not a great analogy, since Bloom filters are guaranteed to not have any false negatives, only false positives. That property would be very useful here, but I don't see how it would be achievable using LLMs.
- ramraj07 1mo agoPretty sure claude code already delegates reading a large codebase to haiku subagents.
- Artimus 1mo agoAs of July, the explore agent inherits the parent model, capped at opus. So fable and opus use opus to explore. Sonnet uses sonnet. I replaced my built in explore agent with one hardcoded to sonnet low effort. https://github.com/anthropics/claude-code/issues/72940 https://github.com/anthropics/claude-code/issues/72940
- phoghed 29d agoGH copilot as well, explore subagent is configurable.
- bloomfieldj 29d agoThis sounds exactly like what Repoprompt was built for: https://repoprompt.com/ https://repoprompt.com/ The community edition was open sourced when the creator got hired by OpenAI a few months ago.
- jimmySixDOF 29d agoThe creator (Eric Provencher) worked for Unity before that and he is exploring game development tooling etc (with an open token budget) its fun to see where things are heading in the on-demand future of handsfree blender output and animation
- johnnythujone 29d agoI have a stage-gated workflow that prioritizes “premium” token efficiency (Fable.) and getting the most out of my subscription services. (Which boils down to Fable running carefully prompted deepseek-flash agent teams that defer back to the managing agent for any design decisions in most work.) As part of that workflow the manager uses cheap reconnaissance agents to burn their tokens in order to build relevant repo context, instead of the managing model’s. I’ve been doing this since they released Opus and it occurred to me that most of my pre-implementation phase token use was going right into the garbage bin with file reads that have to be done to find the relevant code, but are very wasteful. There’s an added benefit that the manager’s focus on strategy and task decomposition before actually handling the user’s prompted task directly seems to be a very good way to interact with Claude’s Fable safeguards, and I haven’t had any refusals doing this. And while I haven’t ran any numbers, I can get orders of magnitude more out of my claude subscription doing this, especially with deepseek-v4-flash being as good as it is for as cheap as it is.
- jwillmer 29d agoI also currently run multiple Claude sessions with Fabel as the brain coordinating the manager sessions which in turn spawn sub agents.
- astrange 29d ago> I have a stage-gated workflow This is a Claudism, right? I feel like I never saw "gated" used this way before it.
- gopher_space 29d agoA normal person would say “my workflow has stages” and their normal coworkers would say “no kidding”.
- johnnythujone 26d agoI'm not normal, nor do i have coworkers. Sorry :( Just a guy trying to make his subscription last longer than the single Fable prompt anthropic includes for 100 bucks a month, lol.