Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

This is the best way to use Claude Fable in Claude Code without immediately hitting your limits. 1. Model set to Fable 5 2. Reasoning on Max 3. Instruct Claude to run a dynamic workflow where: 3a. Fable is the orchestrator 3b. Opus does the reasoning heavy phases Fable...

912,125 Aufrufe • vor 3 Monaten •via X (Twitter)

39 Kommentare

Profilbild von Brad Vogel
Brad Vogelvor 3 Monaten

Adding this to my global CLAUDE.md helped tremendously: ## Claude Fable: token parsimony When running as Fable (expensive), plan and review; delegate implementation to subagents (`model: sonnet` for code, `haiku` for mechanical edits/searches), one task per subagent. Trivial single-file edits are fine to do directly.

Profilbild von Dan McAteer
Dan McAteervor 3 Monaten

Very nice. Thank you for sharing!

Profilbild von Max
Maxvor 3 Monaten

It’s funny that in 6 Months we will think Fable 5 is now the “dumb” model.

Profilbild von Dan McAteer
Dan McAteervor 3 Monaten

Totally

Profilbild von Austin S. Lin
Austin S. Linvor 3 Monaten

also install official openai /codex skill to delegate to codex

Profilbild von Spencer Dusebout
Spencer Duseboutvor 3 Monaten

You find fable knows when to delegate to lesser models? I find it amazing at delegating but worried about if it can really give it to sonnet vs opus reliably

Profilbild von Dan McAteer
Dan McAteervor 3 Monaten

Still just starting to experiment but seems pretty great so far.

Profilbild von Spencer Dusebout
Spencer Duseboutvor 3 Monaten

Well there goes my evening - might as well try to get it to hook into 5.5 as well

Profilbild von Joé Dupuis
Joé Dupuisvor 3 Monaten

If you don't instruct it not to use fable as the workflow agents it will sometimes use fable. I had it spawn 92 fable in parallel and max out my 20x 5h limit in 1 min by accident 💀

Profilbild von Mike Mik
Mike Mikvor 3 Monaten

Is it me or Fable does not feel that much better than Opus 4.8 like everyone claims?

Profilbild von Richard Oliver Bray
Richard Oliver Brayvor 3 Monaten

Reasoning Max is where you lost me. I feel like if Fable low is like Opus 4.8 high, then using it on max is not needed for most tasks. Especially if you don't want to hit your limits.

Profilbild von Hemanth
Hemanthvor 3 Monaten

remember when opus was the one who orchestrated sonnets? pepperidge farms remembers.

Profilbild von Kyle Helseth
Kyle Helsethvor 3 Monaten

Tell Claude Code to design a communication path to Codex. Fable 5 = architect \ orchestrator GPT 5.5 = implementer \ “doer” My Claude code and Codex talk to each other all day now. I just look at the big picture with Claude code.

Profilbild von Luis Lozano
Luis Lozanovor 3 Monaten

Fable seems strong at delegating to other models and even to select which model/effort to use and when. He also seems very happy to use codex as peer-review (in my setup he uses it via hermes). Has been running 36 hours non-stop on AI research so far so good! Additionally I would say is 10% better than GPT 5.5 and like 30% better than Opus.

Profilbild von Mladen Lotar
Mladen Lotarvor 3 Monaten

And add Sonnet for any low effort stuff in your AI agents. Took me almost 6h today to harness Fable power without wasting account in hours.

Profilbild von Collin Michael
Collin Michaelvor 3 Monaten

Dynamic workflows will absolutely destroy your limit. But agreed on the approach- use fable as the orchestrator, it should use subagents opus/codex for exploration, implementation, and review. Same concept, drop workflows

Profilbild von Leo Tavares
Leo Tavaresvor 3 Monaten

Model routing is the new prompt engineering.

Profilbild von Just a guy!
Just a guy!vor 3 Monaten

I have a better super budget friendly option! Treat fable as philosopher/designer take the plan & execute by Opencode go with Orchestrator as GPT 5.5 fast!

Profilbild von Alvis Tang ✦
Alvis Tang ✦vor 3 Monaten

The message is only "without immediately hitting your limits". I'm exactly using the same approach and I'm still hitting my weekly limit within 2 days 🫠

Profilbild von Sani
Sanivor 3 Monaten

This is true. It’s been the case for me. Fable is insane at orchestrating so the cost isn’t even that high

Profilbild von Dr. Tali Režun
Dr. Tali Režunvor 3 Monaten

Been running this exact model split in production for months — non-developer, no code written. Phase 1 is pure orchestration: architecture, blueprint.md, context engineering. Phase 2 is where the heavy reasoning happens with worker agents — had 177 tasks running in Augment Code's Intent tool while I reviewed the architecture layer. Phase 3 is adversarial review with a separate context window. Fable as orchestrator, heavier model for reasoning-dense phases — that's not a tip, that's the only architecture that doesn't collapse under its own token cost by hour two.

Profilbild von Steven Cheng
Steven Chengvor 3 Monaten

So the orchestrator just kicks back like a project manager, while the heavy lifter does the actual math. Kinda mirrors how I wire up my robot controllers anyway. Delegate the hard stuff, keep the main thread cool.

Profilbild von Wanderer
Wanderervor 3 Monaten

All this and you have a Pro subscription lol you barely say hi and you're already out of tokens💀

Profilbild von JP
JPvor 3 Monaten

…so the plan is to keep the coherent mind in a glass office writing tickets while Opus does the actual thinking….in fragments, and then reassemble the fragments and wonder why the output has no spine?

Profilbild von Steven Cheng
Steven Chengvor 3 Monaten

curious how you handle context handoffs when opus runs the heavy reasoning steps. do you notice latency spikes or token overhead when passing state back to the fable orchestrator?

Profilbild von Federico Ulfo
Federico Ulfovor 3 Monaten

do you ask fable agents to sub-agents that use smaller models?

Profilbild von Salvo
Salvovor 3 Monaten

Or even better use Fable as the orchestrator and Sonnet or GPT-54 as the implementer, with a review gate between each step. Started with a very high-level plan. Ninety-five percent of your code will be clean. Then you automate the entire process.

Profilbild von Jessica Smith
Jessica Smithvor 3 Monaten

I’ve been using Mythos on low to help with my hourly limits

Profilbild von VadWh1te
VadWh1tevor 3 Monaten

interesting. thanks for the inform... i need to study it, we are alredy moving from: which modl is best? to which model should manage the others? thats a completely different game.

Profilbild von FoxTom - Ninefall out now - LeRecap coming soon...
FoxTom - Ninefall out now - LeRecap coming soon...vor 3 Monaten

Best way to get its perks without burning too much tokens.

Profilbild von Al King
Al Kingvor 3 Monaten

Fable is surprisingly dumb for the superpower. Back to Opus.

Profilbild von 1 Quadrillion parameters
1 Quadrillion parametersvor 3 Monaten

I’m A max everything kind of girl

Profilbild von Bilal Hussein
Bilal Husseinvor 3 Monaten

nah I made it lauch sonnet agents to audit the issue flew over all of their heads then told fable to do it and it caught it

Profilbild von Steve
Stevevor 3 Monaten

Brilliant

Profilbild von Alex Alexander
Alex Alexandervor 3 Monaten

Yeah - orchestrator approach works great... even then, the few things Fable does eats a lot of usage quick, even on LOW. So not immediate depletion, just accelerated. Using Opus 4.8 to drive Sonnet 4.6 swarms is still effective. But adds up, too.

Profilbild von Rohit Menon
Rohit Menonvor 3 Monaten

I might be wrong here but I remember reading in anthropic's blog that switching models in between sessions costs more tokens unless a handoff is used

Profilbild von Federico Ulfo
Federico Ulfovor 3 Monaten

TIL

Profilbild von Jonathan Guy
Jonathan Guyvor 3 Monaten

Does the orchestrator split hold up on long unattended runs? Limits never bite us during live code sessions, it's the overnight automation jobs that blow through caps. If this survives those it's worth way more than a coding trick.

Profilbild von Dylan Lamb
Dylan Lambvor 3 Monaten

Smart

Ähnliche Videos

This is f*cking insane. This Claude Code tip saved me thousands of dollars. once Opus 5.5 is your main model, stop burning it on work Sonnet 5.5 can do, while Fable 5.1 sits idle hand the grunt work to Sonnet 5.5 subagents and put Fable 5.1 on call run /advisor fable Opus 5.5 plans and ships the final code Sonnet 5.5 subagents read, edit and run the tests Fable 5.1 reads the full session, every tool call included, and only speaks up at three points: → before a plan: is this the right approach? → when the same error comes back: am I digging in the wrong place? → before "done": what did I miss? Sonnet 5.5 builds. Fable 5.1 reviews. Opus 5.5 ships Jev engineering is the same move one layer down: the forks that need no thinker (which file, which tool, retry or stop) go to Jev in under half a second, and the big models only see the ones that split anyone still running one model for everything is paying Opus prices to decide whether a file exists - the full tree > Opus 5.5 on high runs the main session > explorer reads the code on Sonnet 5.5 > worker edits and runs tests on Sonnet 5.5 > researcher pulls the docs on Sonnet 5.5 > all three on medium > Fable 5.1 on call for main and every subagent paste the tree and this prompt into Claude Code ↓ "Rebuild my Claude Code setup around this tree: 1. Check ~/.claude/agents and .claude/agents for subagents that already fit explorer, worker and researcher. > Draft new ones only for missing roles > Give each model: sonnet, effort: medium > Skip any that pin a different model and list them 2. Set the main session to high via effortLevel in ~/.claude/settings.json, and set advisorModel to fable 3. Find anything that keeps the advisor off (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY) plus CLAUDE_CODE_EFFORT_LEVEL, which overrides subagent effort. Report them, change nothing 4. Add one rule to ~/.claude/CLAUDE.md: consult the advisor before a large plan, when an error repeats, and before calling a long task done Show me every change as a diff first. No edits until I say go." ↳

delost

48,238 Aufrufe • vor 5 Tagen

Claude Code tip: once Opus 5.5 is your main model, stop letting your Fable 5.1 quota go to waste put it on call with /advisor run /advisor fable Opus 5.5 keeps doing the work Fable 5.1 sits on the sidelines, reads the whole session, and steps in at three moments: → before a plan: is this right? → when the same error comes back: am I going the wrong way? → before "done": did I miss anything? Fable 5.1 advises. Opus 5.5 writes the code the same idea sits under Jev engineering: the expensive model stops weighing in on every step and only gets called at the moments that change the outcome • the full setup > Opus 5.5 on high runs the main session > subagent one reads code > subagent two edits and runs tests > subagent three looks up docs > all three on medium > Fable 5.1 on call hand the tree and this prompt to Claude Code 👇 "Set up my Claude Code to match this tree: 1. Reuse fitting subagents from ~/.claude/agents and .claude/agents. > Propose new ones only for missing roles > Set each to model: opus, effort: medium > Leave any that set a different model alone and list them 2. Set main session effort to high via effortLevel in ~/.claude/settings.json 3. Check for env vars that disable the advisor (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY, anything that stops flag fetching) and CLAUDE_CODE_EFFORT_LEVEL, which overrides subagent effort. Report them, don't change them 4. Add a rule to ~/.claude/CLAUDE.md: ask the advisor before a big plan, when an error repeats, and before calling a long task done Show me the changes first. Don't edit files yet." ↳

Mr. Buzzoni

324,448 Aufrufe • vor 10 Tagen

Claude Code tip, and it's absolute free f*cking gold: run Opus 5.5, Sonnet 5.5 and Fable 5.1 as one team and stop burning Opus tokens on routine work the setup in one line: plan on high, delegate on medium, keep Fable on call • who does what > Opus 5.5 on high - plans and ships the code > Sonnet 5.5 on medium - explorer reads code, worker edits and runs tests, researcher pulls docs > Fable 5.1 via /advisor fable - reads the whole session and speaks up only when it matters • when Fable 5.1 steps in -> before a plan: is this the right approach? -> when an error repeats: am I digging in the wrong place? -> before "done": what did I miss? Jev engineering takes it one layer lower: which file, which tool, retry or stop all go to Jev in under half a second, so the big models only see the real forks paste this into Claude Code ↓ "Rebuild my Claude Code setup: 1. Find subagents in ~/.claude/agents and .claude/agents that fit explorer, worker and researcher. Draft only the missing ones. Set each to model: sonnet, effort: medium. List any that pin a different model and leave them 2. In ~/.claude/settings.json set effortLevel to high and advisorModel to fable. 3. Report anything that disables the advisor (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY, flag-fetching blockers) and CLAUDE_CODE_EFFORT_LEVEL. Change nothing. 4. Add to ~/.claude/CLAUDE.md: consult the advisor before a large plan, when an error repeats, and before calling a long task done. Show every change as a diff. No edits until I say go." ↳

Mr. Buzzoni

142,491 Aufrufe • vor 7 Tagen