Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Opus 5.5 is quietly the best value model Anthropic has shipped. Anthropic just published the official playbook for Opus 5.5, written by Addy Osmani, who spent years on Google's Chrome team. 40% cheaper than Opus 5. Cache reads 60% cheaper. Fable 5.1 level on most tasks. Runs for hours...

11,518 görüntüleme • 4 gün önce •via X (Twitter)

12 Yorum

Hussain Hashim | Building SundayBack profil fotoğrafı
Hussain Hashim | Building SundayBack4 gün önce

@DamiDefi interesting move. wonder how Opus 5.5 handles edge cases compared to Fable 5.1? those usually trip up similar models.

Macro Bombastic profil fotoğrafı
Macro Bombastic4 gün önce

40% cheaper and runs for hours, yeah that changes things tbh

Dami-Defi profil fotoğrafı
Dami-Defi4 gün önce

The best deal fr

8lends profil fotoğrafı
8lends4 gün önce

the pricing shift here actually makes this model way more competitive theyre playing the value game right not just pushing new versions

Dami-Defi profil fotoğrafı
Dami-Defi4 gün önce

Cutting cost, but keeping the value

Pate profil fotoğrafı
Pate4 gün önce

The number I'd want is cost per finished task, not per token. A cheaper model that needs two extra passes isn't cheaper.

AGTP profil fotoğrafı
AGTP4 gün önce

It really is a game changer for value, but many people are still prompting it like older models and losing money. We actually went deeper on this here:

The AI Therapist profil fotoğrafı
The AI Therapist4 gün önce

Opus has always been the reliable workhorse. 40% cheaper? That’s not just value; that’s permission to think bigger without burning cash. A smart mom move in silicon form. Addy knows optimization well.

Divi Stacker profil fotoğrafı
Divi Stacker4 gün önce

Anthropic quietly cooked with this one

Dami-Defi profil fotoğrafı
Dami-Defi4 gün önce

No lies

Sujal profil fotoğrafı
Sujal4 gün önce

cache reads only get cheap if the prefix is byte identical, a timestamp near the top and you pay the write every single time

Adrian profil fotoğrafı
Adrian4 gün önce

fable level performance is impressive 🔥

Benzer Videolar

How to cut your AI bill by 60% switching to Opus 5.5: Most people will swap the model name, save 24%, and stop there. The other 37% is sitting in your settings. Here's the real before and after on an example agent workload. One month: 100M input tokens (80M cached reads, 5M cache writes, 15M uncached) 10M output tokens BEFORE: Opus 5 Cache reads: 80M × $0.50 = $40 Cache writes: 5M × $6.25 = $31.25 Uncached input: 15M × $5 = $75 Output: 10M × $25 = $250 Total: $396.25 STEP 1. Just switch the model. Same tokens, new prices. Cache reads $0.20. Writes $5. Input $4. Output $20. $16 + $25 + $60 + $200 = $301 24% cheaper. That's what everyone's screenshotting. STEP 2. Let it write less. Box measured Opus 5.5 at 40% less verbose with no drop in accuracy. 10M output tokens becomes 6M. Output drops from $200 to $120. Total: $221. Now you're at 44% off. STEP 3. Stop running every call at max effort. Thinking can't be switched off on 5.5 anymore, so effort is your lever. Classifying, routing, formatting, summarizing? Drop the effort. Save the high settings for the calls that actually reason. Say that trims output another 25%, 6M down to 4.5M. Output: $90. Total: $191. 52% off. STEP 4. Cache the stuff you keep resending. This is the one nobody does. Cache reads went from $0.50 to $0.20. That's 60% off the cheapest line on your bill. Move your system prompt, tool definitions and repo context into the cache. Uncached input drops from 15M to 5M, cache reads go up to 90M. $18 + $25 + $20 + $90 = $153 AFTER: $153 From $396.25. 61% cheaper. Same work. The model gave you 24%. You gave yourself the other 37%. Anthropic's own number is 40% cheaper than Opus 5 on typical workloads. Your number depends on how much of your bill is output and how much of your prompt you're resending uncached every call. So check those two first. Bookmark this for when you migrate. follow CyrilXBT

CyrilXBT

15,803 görüntüleme • 7 gün önce

fable 5.1 vs fable 5 vs opus 5 – three lord of the rings landmarks, built in 3d from one image the setup: one reference image per scene, one html file per build, everything procedural – no meshes, no textures, no image files, nothing past Three.js from a cdn. each model reads the picture, writes its own prompt from it, then builds to that prompt in the same turn. three named camera shots per scene on keys 1/2/3, so it can be screen-recorded. run through OpenRouter tasks: 1. bag end – hobbiton from two frames, outside and in. the round green door has to open onto the room you are standing in 2. barad-dûr – the tower and orodruin from one film still. the eye has to move and track the camera, the volcano erupts on a cycle, the clouds never stop 3. rivendell – jerry vanderstelt's painting. sun shafts that shimmer, water that falls without a break, trees that sway on a gust models: Anthropic fable 5.1, fable 5, opus 5 total cost, three builds #1 fable 5 – $14.97 #2 opus 5 – $18.53 #3 fable 5.1 – $22.38 wall clock, three builds #1 fable 5 – 38m #2 fable 5.1 – 92m #3 opus 5 – 122m output tokens #1 fable 5 – 298,592 #2 fable 5.1 – 439,435 #3 opus 5 – 724,418 lines of code shipped #1 fable 5 – 2,885 #2 fable 5.1 – 4,021 #3 opus 5 – 5,161 biggest single build, lines #1 opus 5, bag end – 2,410 #2 fable 5.1, barad-dûr – 1,375 #3 fable 5, bag end – 1,319 observations: • fable 5.1 is the only model that furnished the bag end interior – a live fire, panelling, books on the floor, leaded diamond windows, against fable 5's flat color and opus's dark tunnel. the round door outside opens onto that room, the hard part of the brief • what it costs is thinking room. the 128k output ceiling is a thinking budget in disguise: fable 5.1 burned 102,116 of it on reasoning and hit the wall mid-file. opus spent 109,241 and hit the same wall. fable 5 spent 61,240 and finished bag end in one call – the only one that did • fable 5.1's first pass is not the finished thing. its barad-dûr came back with three defects you only catch by looking at it – nothing a read of the code would have flagged • it is the best of the three at being corrected. handed a plain list of what was wrong, it returned 32 targeted patches over two rounds, every one applied first try, and it worked out one of the causes itself instead of guessing at constants conclusion: nine scenes, 12,067 lines and 1.46m output tokens for $55.88 all in – and the cheapest model was also the fastest, by 3.2x! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

18,509 görüntüleme • 27 gün önce