正在加载视频...

视频加载失败

This is the best way to use Claude Fable in Claude Code without immediately hitting your limits. 1. Model set to Fable 5 2. Reasoning on Max 3. Instruct Claude to run a dynamic workflow where: 3a. Fable is the orchestrator 3b. Opus does the reasoning heavy phases Fable...

912,125 次观看 • 3 个月前 •via X (Twitter)

39 条评论

Brad Vogel 的头像
Brad Vogel3 个月前

Adding this to my global CLAUDE.md helped tremendously: ## Claude Fable: token parsimony When running as Fable (expensive), plan and review; delegate implementation to subagents (`model: sonnet` for code, `haiku` for mechanical edits/searches), one task per subagent. Trivial single-file edits are fine to do directly.

Dan McAteer 的头像
Dan McAteer3 个月前

Very nice. Thank you for sharing!

Max 的头像
Max3 个月前

It’s funny that in 6 Months we will think Fable 5 is now the “dumb” model.

Dan McAteer 的头像
Dan McAteer3 个月前

Totally

Austin S. Lin 的头像
Austin S. Lin3 个月前

also install official openai /codex skill to delegate to codex

Spencer Dusebout 的头像
Spencer Dusebout3 个月前

You find fable knows when to delegate to lesser models? I find it amazing at delegating but worried about if it can really give it to sonnet vs opus reliably

Dan McAteer 的头像
Dan McAteer3 个月前

Still just starting to experiment but seems pretty great so far.

Spencer Dusebout 的头像
Spencer Dusebout3 个月前

Well there goes my evening - might as well try to get it to hook into 5.5 as well

Joé Dupuis 的头像
Joé Dupuis3 个月前

If you don't instruct it not to use fable as the workflow agents it will sometimes use fable. I had it spawn 92 fable in parallel and max out my 20x 5h limit in 1 min by accident 💀

Mike Mik 的头像
Mike Mik3 个月前

Is it me or Fable does not feel that much better than Opus 4.8 like everyone claims?

Richard Oliver Bray 的头像
Richard Oliver Bray3 个月前

Reasoning Max is where you lost me. I feel like if Fable low is like Opus 4.8 high, then using it on max is not needed for most tasks. Especially if you don't want to hit your limits.

Hemanth 的头像
Hemanth3 个月前

remember when opus was the one who orchestrated sonnets? pepperidge farms remembers.

Kyle Helseth 的头像
Kyle Helseth3 个月前

Tell Claude Code to design a communication path to Codex. Fable 5 = architect \ orchestrator GPT 5.5 = implementer \ “doer” My Claude code and Codex talk to each other all day now. I just look at the big picture with Claude code.

Luis Lozano 的头像
Luis Lozano3 个月前

Fable seems strong at delegating to other models and even to select which model/effort to use and when. He also seems very happy to use codex as peer-review (in my setup he uses it via hermes). Has been running 36 hours non-stop on AI research so far so good! Additionally I would say is 10% better than GPT 5.5 and like 30% better than Opus.

Mladen Lotar 的头像
Mladen Lotar3 个月前

And add Sonnet for any low effort stuff in your AI agents. Took me almost 6h today to harness Fable power without wasting account in hours.

Collin Michael 的头像
Collin Michael3 个月前

Dynamic workflows will absolutely destroy your limit. But agreed on the approach- use fable as the orchestrator, it should use subagents opus/codex for exploration, implementation, and review. Same concept, drop workflows

Leo Tavares 的头像
Leo Tavares3 个月前

Model routing is the new prompt engineering.

Just a guy! 的头像
Just a guy!3 个月前

I have a better super budget friendly option! Treat fable as philosopher/designer take the plan & execute by Opencode go with Orchestrator as GPT 5.5 fast!

Alvis Tang ✦ 的头像
Alvis Tang ✦3 个月前

The message is only "without immediately hitting your limits". I'm exactly using the same approach and I'm still hitting my weekly limit within 2 days 🫠

Sani 的头像
Sani3 个月前

This is true. It’s been the case for me. Fable is insane at orchestrating so the cost isn’t even that high

Dr. Tali Režun 的头像
Dr. Tali Režun3 个月前

Been running this exact model split in production for months — non-developer, no code written. Phase 1 is pure orchestration: architecture, blueprint.md, context engineering. Phase 2 is where the heavy reasoning happens with worker agents — had 177 tasks running in Augment Code's Intent tool while I reviewed the architecture layer. Phase 3 is adversarial review with a separate context window. Fable as orchestrator, heavier model for reasoning-dense phases — that's not a tip, that's the only architecture that doesn't collapse under its own token cost by hour two.

Steven Cheng 的头像
Steven Cheng3 个月前

So the orchestrator just kicks back like a project manager, while the heavy lifter does the actual math. Kinda mirrors how I wire up my robot controllers anyway. Delegate the hard stuff, keep the main thread cool.

Wanderer 的头像
Wanderer3 个月前

All this and you have a Pro subscription lol you barely say hi and you're already out of tokens💀

JP 的头像
JP3 个月前

…so the plan is to keep the coherent mind in a glass office writing tickets while Opus does the actual thinking….in fragments, and then reassemble the fragments and wonder why the output has no spine?

Steven Cheng 的头像
Steven Cheng3 个月前

curious how you handle context handoffs when opus runs the heavy reasoning steps. do you notice latency spikes or token overhead when passing state back to the fable orchestrator?

Federico Ulfo 的头像
Federico Ulfo3 个月前

do you ask fable agents to sub-agents that use smaller models?

Salvo 的头像
Salvo3 个月前

Or even better use Fable as the orchestrator and Sonnet or GPT-54 as the implementer, with a review gate between each step. Started with a very high-level plan. Ninety-five percent of your code will be clean. Then you automate the entire process.

Jessica Smith 的头像
Jessica Smith3 个月前

I’ve been using Mythos on low to help with my hourly limits

VadWh1te 的头像
VadWh1te3 个月前

interesting. thanks for the inform... i need to study it, we are alredy moving from: which modl is best? to which model should manage the others? thats a completely different game.

FoxTom - Ninefall out now - LeRecap coming soon... 的头像
FoxTom - Ninefall out now - LeRecap coming soon...3 个月前

Best way to get its perks without burning too much tokens.

Al King 的头像
Al King3 个月前

Fable is surprisingly dumb for the superpower. Back to Opus.

1 Quadrillion parameters 的头像
1 Quadrillion parameters3 个月前

I’m A max everything kind of girl

Bilal Hussein 的头像
Bilal Hussein3 个月前

nah I made it lauch sonnet agents to audit the issue flew over all of their heads then told fable to do it and it caught it

Steve 的头像
Steve3 个月前

Brilliant

Alex Alexander 的头像
Alex Alexander3 个月前

Yeah - orchestrator approach works great... even then, the few things Fable does eats a lot of usage quick, even on LOW. So not immediate depletion, just accelerated. Using Opus 4.8 to drive Sonnet 4.6 swarms is still effective. But adds up, too.

Rohit Menon 的头像
Rohit Menon3 个月前

I might be wrong here but I remember reading in anthropic's blog that switching models in between sessions costs more tokens unless a handoff is used

Federico Ulfo 的头像
Federico Ulfo3 个月前

TIL

Jonathan Guy 的头像
Jonathan Guy3 个月前

Does the orchestrator split hold up on long unattended runs? Limits never bite us during live code sessions, it's the overnight automation jobs that blow through caps. If this survives those it's worth way more than a coding trick.

Dylan Lamb 的头像
Dylan Lamb3 个月前

Smart

相关视频

This is f*cking insane. This Claude Code tip saved me thousands of dollars. once Opus 5.5 is your main model, stop burning it on work Sonnet 5.5 can do, while Fable 5.1 sits idle hand the grunt work to Sonnet 5.5 subagents and put Fable 5.1 on call run /advisor fable Opus 5.5 plans and ships the final code Sonnet 5.5 subagents read, edit and run the tests Fable 5.1 reads the full session, every tool call included, and only speaks up at three points: → before a plan: is this the right approach? → when the same error comes back: am I digging in the wrong place? → before "done": what did I miss? Sonnet 5.5 builds. Fable 5.1 reviews. Opus 5.5 ships Jev engineering is the same move one layer down: the forks that need no thinker (which file, which tool, retry or stop) go to Jev in under half a second, and the big models only see the ones that split anyone still running one model for everything is paying Opus prices to decide whether a file exists - the full tree > Opus 5.5 on high runs the main session > explorer reads the code on Sonnet 5.5 > worker edits and runs tests on Sonnet 5.5 > researcher pulls the docs on Sonnet 5.5 > all three on medium > Fable 5.1 on call for main and every subagent paste the tree and this prompt into Claude Code ↓ "Rebuild my Claude Code setup around this tree: 1. Check ~/.claude/agents and .claude/agents for subagents that already fit explorer, worker and researcher. > Draft new ones only for missing roles > Give each model: sonnet, effort: medium > Skip any that pin a different model and list them 2. Set the main session to high via effortLevel in ~/.claude/settings.json, and set advisorModel to fable 3. Find anything that keeps the advisor off (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY) plus CLAUDE_CODE_EFFORT_LEVEL, which overrides subagent effort. Report them, change nothing 4. Add one rule to ~/.claude/CLAUDE.md: consult the advisor before a large plan, when an error repeats, and before calling a long task done Show me every change as a diff first. No edits until I say go." ↳

delost

48,238 次观看 • 5 天前

Claude Code tip: once Opus 5.5 is your main model, stop letting your Fable 5.1 quota go to waste put it on call with /advisor run /advisor fable Opus 5.5 keeps doing the work Fable 5.1 sits on the sidelines, reads the whole session, and steps in at three moments: → before a plan: is this right? → when the same error comes back: am I going the wrong way? → before "done": did I miss anything? Fable 5.1 advises. Opus 5.5 writes the code the same idea sits under Jev engineering: the expensive model stops weighing in on every step and only gets called at the moments that change the outcome • the full setup > Opus 5.5 on high runs the main session > subagent one reads code > subagent two edits and runs tests > subagent three looks up docs > all three on medium > Fable 5.1 on call hand the tree and this prompt to Claude Code 👇 "Set up my Claude Code to match this tree: 1. Reuse fitting subagents from ~/.claude/agents and .claude/agents. > Propose new ones only for missing roles > Set each to model: opus, effort: medium > Leave any that set a different model alone and list them 2. Set main session effort to high via effortLevel in ~/.claude/settings.json 3. Check for env vars that disable the advisor (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY, anything that stops flag fetching) and CLAUDE_CODE_EFFORT_LEVEL, which overrides subagent effort. Report them, don't change them 4. Add a rule to ~/.claude/CLAUDE.md: ask the advisor before a big plan, when an error repeats, and before calling a long task done Show me the changes first. Don't edit files yet." ↳

Mr. Buzzoni

324,448 次观看 • 10 天前

Claude Code tip, and it's absolute free f*cking gold: run Opus 5.5, Sonnet 5.5 and Fable 5.1 as one team and stop burning Opus tokens on routine work the setup in one line: plan on high, delegate on medium, keep Fable on call • who does what > Opus 5.5 on high - plans and ships the code > Sonnet 5.5 on medium - explorer reads code, worker edits and runs tests, researcher pulls docs > Fable 5.1 via /advisor fable - reads the whole session and speaks up only when it matters • when Fable 5.1 steps in -> before a plan: is this the right approach? -> when an error repeats: am I digging in the wrong place? -> before "done": what did I miss? Jev engineering takes it one layer lower: which file, which tool, retry or stop all go to Jev in under half a second, so the big models only see the real forks paste this into Claude Code ↓ "Rebuild my Claude Code setup: 1. Find subagents in ~/.claude/agents and .claude/agents that fit explorer, worker and researcher. Draft only the missing ones. Set each to model: sonnet, effort: medium. List any that pin a different model and leave them 2. In ~/.claude/settings.json set effortLevel to high and advisorModel to fable. 3. Report anything that disables the advisor (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY, flag-fetching blockers) and CLAUDE_CODE_EFFORT_LEVEL. Change nothing. 4. Add to ~/.claude/CLAUDE.md: consult the advisor before a large plan, when an error repeats, and before calling a long task done. Show every change as a diff. No edits until I say go." ↳

Mr. Buzzoni

142,491 次观看 • 7 天前