Loading video...

Video Failed to Load

Go Home

Claude Code tip: once Opus 5.5 is your main model, stop letting your Fable 5.1 quota go to waste put it on call with /advisor run /advisor fable Opus 5.5 keeps doing the work Fable 5.1 sits on the sidelines, reads the whole session, and steps in at three...

319,285 views • 2 days ago •via X (Twitter)

34 Comments

Suraj Mandal's profile picture
Suraj Mandal1 day ago

Fable draws from your same weekly quota, I'd rather use all of it using opus since it feels better to use than fable right now usage wise.

TheEpTic's profile picture
TheEpTic1 day ago

Doesn’t fable eat the weekly too? I hit 100% on weekly the other day and tried fable to see if it worked but it didn’t

Freedom22's profile picture
Freedom221 day ago

Fable doesn't need to step in, it is inferior in everything to opus 5.5

Max Reid's profile picture
Max Reid2 days ago

I have been letting Fable plan and direct Opus until 5.5 came along. Opus is really quick now. That makes it easy to forget about Fable. The long waiting times were my biggest complaint of using Fable + subagents.

Lay's profile picture
Lay2 days ago

the repeat-error trigger is the part i needed. my overnight loops keep rediscovering the same broken fix at 3am and waking me up for it

The AI Therapist's profile picture
The AI Therapist2 days ago

J'aime cette astuce : Opus 5.5 pilote le tout pendant que Fable surveille les détails pour toi. Malin !

Jeremy Longshore's profile picture
Jeremy Longshore1 day ago

API key required.

john🎯 aka Mr. Producer's profile picture
john🎯 aka Mr. Producer1 day ago

Opus is better. Cleaner. Hardly touching fable rn.

riVeN's profile picture
riVeN1 day ago

solid. just note the docs say advisor timing is model-driven, so that claude.md rule is doing the real work.

Roshni's profile picture
Roshni1 day ago

Maximizing multi-agent delegation like this makes development so much more efficient.

Final Miro's profile picture
Final Miro1 day ago

ok buddy

Max Bevza's profile picture
Max Bevza1 day ago

This is basically model specialization without paying the expensive inference cost on every step

Steven | nevetS's profile picture
Steven | nevetS1 day ago

AI backseating AI

山川OK50饭's profile picture
山川OK50饭1 day ago

把好钢都用在刀刃上这招确实高明

David Cumps's profile picture
David Cumps1 day ago

you're still burning weekly credits though

Alex's profile picture
Alex2 days ago

Using the second model as a reader that only steps in at decision points is a cleaner split than running two agents on the same task. The hard part is keeping those moments few. If the advisor starts commenting on every step, the quota win disappears.

Fede's profile picture
Fede1 day ago

@grok come dovrei impostare questi subagenti di fable 5.1?

Ratel's profile picture
Ratel1 day ago

interesting that you have it on advisor the general version is separating available capability from active context. your system can have access to a much larger surface than what needs to participate in every turn.

YunCuntu | Video & Photo Downloader's profile picture
YunCuntu | Video & Photo Downloader1 day ago

Having a second model only chime in at decision points feels like a solid way to keep costs sane without losing the sanity check.

m-check1B's profile picture
m-check1B1 day ago

bravo but not bravo

小北繁70|BG's profile picture
小北繁70|BG1 day ago

这波算力分配省心又高效

Moamen's profile picture
Moamen2 days ago

Same idea I landed on without the advisor command. Cheap model stays in the session. Expensive model only gets the decision points

安叫兽|Bird🕊️ 🔶 BNB's profile picture
安叫兽|Bird🕊️ 🔶 BNB1 day ago

这个搭配还挺实用,就是怕旁观的比干活的还费配额。

Vibes McDeploy's profile picture
Vibes McDeploy1 day ago

the advisor kept asking "is this right?" so i muted it

晚晚's profile picture
晚晚1 day ago

这套分工挺省脑子,像给主力配了个场外教练

Karol's profile picture
Karol1 day ago

That looks interesting defenitly would love to test it out, but I'm sure my claude would mess it up either way

Hrundel75 🐷's profile picture
Hrundel75 🐷1 day ago

this post is pure gold

def not sofia's profile picture
def not sofia1 day ago

This is so true

Arbaz's profile picture
Arbaz1 day ago

advisor is just polite for don't burn opus

Charlie Deist's profile picture
Charlie Deist1 day ago

this seems smart but also like it's overindexing on something. I'm not sure what.

AIガチ勢オジ|40代×Claude Code's profile picture
AIガチ勢オジ|40代×Claude Code1 day ago

メインモデルとアドバイザーを分けて、クォータを無駄にしない運用は本当に参考になります。3つの介入タイミングも試してみます。ありがとうございます!

helicerat's profile picture
helicerat1 day ago

woah bro, setting this up right now

JoshuaLP's profile picture
JoshuaLP1 day ago

Don’t do this, it is not needed

notgwapo's profile picture
notgwapo1 day ago

sounds like a decent setup. gotta see if fable 5.1 throws any curveballs while opus keeps grinding. 🤔

Related Videos

Claude Code tip: once Opus 5.5 is your main model, stop leaving Fable 5.1 sitting idle put it on call with /advisor run /advisor fable Opus 5.5 keeps writing the code Fable 5.1 reads the full session, every tool call included, and only speaks up at three points: → before a plan: is this the right approach? → when the same error comes back: am I digging in the wrong place? → before "done": what did I miss? Fable 5.1 reviews. Opus 5.5 ships Jev engineering is the same move one layer down: the forks that need no thinker (which file, which tool, retry or stop) go to Jev in under half a second, and the big model only sees the ones that split - the full tree > Opus 5.5 on high runs the main session > explorer reads the code > worker edits and runs tests > researcher pulls the docs > all three on medium > Fable 5.1 on call as the advisor paste the tree and this prompt into Claude Code ↓ "Rebuild my Claude Code setup around this tree: 1. Check ~/.claude/agents and .claude/agents for subagents that already fit explorer, worker and researcher. > Draft new ones only for missing roles > Give each model: opus, effort: medium > Skip any that pin a different model and list them 2. Set the main session to high via effortLevel in ~/.claude/settings.json, and set advisorModel to fable 3. Find anything that keeps the advisor off (CLAUDE_CODE_DISABLE_ADVISOR_TOOL, DISABLE_TELEMETRY, any variable that stops feature-flag fetching) plus CLAUDE_CODE_EFFORT_LEVEL, which overrides subagent effort. Report them, change nothing 4. Add one rule to ~/.claude/CLAUDE.md: consult the advisor before a large plan, when an error repeats, and before calling a long task done Show me every change as a diff first. No edits until I say go." ↳

delost

647,583 views • 1 day ago

Jev + Opus 5.5: Anthropic's new model beats GPT-6 Astra for 1/5 the cost, and 4 API changes will 400 your agent before it writes a single line I pulled these 10 steps from the migration docs so you don't learn them in production step 1 → $4 / $20 per 1M. Opus 5 was $5 / $25. cache reads dropped from $0.50 to $0.20 step 2 → 66.4% on Terminal-Bench 4.0 vs GPT-6 Astra 57.9% and Opus 5 52.3%. +14.1 points in one release, and on FrontierCode it beats Astra at default effort for 1/5 the cost step 3 → thinking can't be turned off anymore. send thinking: disabled and you get a 400. drop the field, set effort step 4 → tool_choice any and tool are gone. 400. switch to auto + strict step 5 → edit anything above a thinking block and the request dies. append only, or opt into drop_block step 6 → computer_20251124 is dead on the API. 400. move to computer_toolset_20260801 step 7 → the quiet one: default effort fell from high to medium. your agent thinks less than you set it up to and nothing tells you step 8 → hop Opus 5.5 → Sonnet 5 → Opus 5.5 and you pay 4.36 instead of 3.32. +31%, the cache dies and Sonnet can't read Opus's reasoning step 9 → change effort at the top of the request and the cache is gone. Jev sets it per message and the cache stays step 10 → switch fast - standard mid-session and it's a full cache miss. Jev picks speed once, on turn one one model, three knobs, zero 400s. that is Jev + Opus 5.5 send this to your Claude Code before you touch the model ID, then read my full Jev deep dive in the article below ↓

Carnage

16,674 views • 7 days ago

fable 5.1 vs fable 5 vs opus 5 – three lord of the rings landmarks, built in 3d from one image the setup: one reference image per scene, one html file per build, everything procedural – no meshes, no textures, no image files, nothing past Three.js from a cdn. each model reads the picture, writes its own prompt from it, then builds to that prompt in the same turn. three named camera shots per scene on keys 1/2/3, so it can be screen-recorded. run through OpenRouter tasks: 1. bag end – hobbiton from two frames, outside and in. the round green door has to open onto the room you are standing in 2. barad-dûr – the tower and orodruin from one film still. the eye has to move and track the camera, the volcano erupts on a cycle, the clouds never stop 3. rivendell – jerry vanderstelt's painting. sun shafts that shimmer, water that falls without a break, trees that sway on a gust models: Anthropic fable 5.1, fable 5, opus 5 total cost, three builds #1 fable 5 – $14.97 #2 opus 5 – $18.53 #3 fable 5.1 – $22.38 wall clock, three builds #1 fable 5 – 38m #2 fable 5.1 – 92m #3 opus 5 – 122m output tokens #1 fable 5 – 298,592 #2 fable 5.1 – 439,435 #3 opus 5 – 724,418 lines of code shipped #1 fable 5 – 2,885 #2 fable 5.1 – 4,021 #3 opus 5 – 5,161 biggest single build, lines #1 opus 5, bag end – 2,410 #2 fable 5.1, barad-dûr – 1,375 #3 fable 5, bag end – 1,319 observations: • fable 5.1 is the only model that furnished the bag end interior – a live fire, panelling, books on the floor, leaded diamond windows, against fable 5's flat color and opus's dark tunnel. the round door outside opens onto that room, the hard part of the brief • what it costs is thinking room. the 128k output ceiling is a thinking budget in disguise: fable 5.1 burned 102,116 of it on reasoning and hit the wall mid-file. opus spent 109,241 and hit the same wall. fable 5 spent 61,240 and finished bag end in one call – the only one that did • fable 5.1's first pass is not the finished thing. its barad-dûr came back with three defects you only catch by looking at it – nothing a read of the code would have flagged • it is the best of the three at being corrected. handed a plain list of what was wrong, it returned 32 targeted patches over two rounds, every one applied first try, and it worked out one of the causes itself instead of guessing at constants conclusion: nine scenes, 12,067 lines and 1.46m output tokens for $55.88 all in – and the cheapest model was also the fastest, by 3.2x! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

18,509 views • 27 days ago

i finally mastered how to maximise my opus 5.5 usage limits... the trick: let jev choose which subagent gets each task and how much effort it should use. [here’s how i’d wire it:] claude breaks the project into tasks. jev selects from predefined worker profiles. claude applies the selected settings and dispatches the work. → main session, medium: clarify the requirements, define what “done” looks like, and prepare the tasks → builder, low: small, clearly defined tasks with existing examples or patterns → builder, medium: tasks that connect multiple parts or need decisions within the approved plan → verifier, high: check requirements, probe edge cases, and report problems for the builder to fix jev gets the task’s scope, what’s uncertain, and the consequences of failure. it chooses from the profiles allowed for that task. your approval checkpoints stay in place. paste this into your next planning session: “use opus 5.5 with jev selecting the worker and effort profile for each task. first, check that a working jev integration is available and that this environment supports separate effort settings for subagents. check for configuration or environment overrides that could prevent those settings from taking effect. if anything is missing, explain what needs wiring before proceeding. break my request into tasks with clear ownership, dependencies, relevant context, and acceptance checks. keep small related tasks together when a separate subagent would add unnecessary overhead. keep the main session at medium effort. offer jev these worker profiles: builder at low effort for small, clearly defined tasks using existing patterns; builder at medium effort for tasks that connect multiple parts or require decisions within the approved plan; verifier at high effort for checking requirements and edge cases. give jev each task’s scope, uncertainties, dependencies, and consequences of failure. only offer profiles appropriate to the current stage. validate its selection before dispatching. use the actual jev integration; don’t simulate its decisions. if it abstains or returns an invalid choice, stop that handoff and ask me. show me the task plan and proposed assignments before starting. after approval, launch the selected workers with their assigned effort settings, relevant context, file ownership, and completion checks. let me review the result before verification. the verifier may add tests but must leave implementation code unchanged. have it report what passed, what failed, and what remains uncertain. send implementation fixes back to the builder, then recheck the affected parts. if a task repeatedly fails, examine the requirements and approach before increasing effort. report available total usage, including jev calls, worker calls, retries, and verification. don’t invent missing data. compare similar completed tasks before claiming savings.” steal this 👇

Avid

31,355 views • 4 days ago