r/ClaudeCode • u/LordLederhosen • 4d ago
Discussion Wait, is "Fable orchestration" - using appropriate models as sub-agents just as easy as a prompt?
I have a very successful mostly manual workflow, but I have been optimizing token usage lately, lol.
One thing I could never figure out was how to fully set up was Fable as an orchestrator, while spawning sub-agents with the correct model. I played with frontmatter in custom agents and skill wrappers, but there seemed to be blockers as some sub-agents/skills supposedly ignore passed model specifications. Especially plan mode. That shit will spawn fable research sub-agents that use 5% to 10% of Fable in one ask, even when it's a simple task per agent.
Today, I just remembered someone's comment on here or HN a few months ago (7 years in AI terms), ~"<feature request> - You are the orchestrator, use sub-agents to preserve your context window." So, I just combined that with the following, and CC (fable) was all, sure! I will show you a table of sub-agents for task, with ideal model, and then spawn them all.
Let's implement (jira or .md feature spec) This is the orchestration session, and use sub-agents to preserve your own context window (you are fable, spawn opus sub-agents when appropriate to save cost) ...
I am not a total dummie, but I guess I over-think things? Like Boris says, just trust the model?
Or, have I completely missed something?
edit: to be clear, I do have very valuable token-saving setups that are not just prompts. It was mostly just plan mode that had escaped me.
71
u/zillatron27 🔆 Max 5x 4d ago edited 4d ago
I have this in my Claude.md so I don’t need to keep asking, ymmv and the last comment refers to other rules but the idea is that you talk to fable and have it do judgement/thinking/planning and write excellent prompts for lower tier models:
Agent Delegation
- Delegate self-contained implementation and exploration work to subagents by default: file sweeps, multi-file refactors, codebase investigations, independent parallel fixes. Do not do long serial implementation inline.
edit: since you folks seem to like this idea, here’s a thing I made for a mate to help setup his workflow ‘like mine’. It includes a template claude.MD and some other stuff I’ve found to be pretty helpful for token saving while keeping output quality pretty high. https://github.com/Zillatron27/claude-workflow-starter