正在加载视频...

视频加载失败

Claude Fable 5 changed how we work on the Claude Code team day to day. We used to verify that Claude did the work right. Now we verify that it's doing the right work. Here’s the 3 biggest changes:

1,042,226 次观看 • 4 个月前 •via X (Twitter)

32 条评论

Hunter Bertoson 的头像
Hunter Bertoson4 个月前

I’m Putting it to the test.

Layton Gott 的头像
Layton Gott4 个月前

Why can't I see it in the desktop app yet?! I have no patience lol.

Nova 的头像
Nova4 个月前

Anthropic right now

Abdallah 的头像
Abdallah4 个月前

It's refusing to scan our apps for security vulnerabilities sadly

Mr. Cool 的头像
Mr. Cool4 个月前

Please do text instead of video. Easier to consume on the go

Brian De Souza 的头像
Brian De Souza4 个月前

moving form "checking if the work is right" to "checking if it's doing the right work". the ideas guy era is upon us

Amit Yadav 的头像
Amit Yadav4 个月前

Everyone's talking about Claude Fable 5's benchmarks. I'm more interested in the workflows that won't exist a year from now because of it.

Matt Matheus 的头像
Matt Matheus4 个月前

Fable as planner and code reviewer is about all we can afford.

Andrzej Shukevich 的头像
Andrzej Shukevich4 个月前

The time has come…

Mücahit 的头像
Mücahit4 个月前

Bırak kolpayı limitleri sıfırla amk

Mohamad Al-Zawahreh 的头像
Mohamad Al-Zawahreh4 个月前

It doesn't even work. It reroutes all tasks. An excuse to get users conversation history If they - while annoyed submit feedback. The sabotage of anyone doing AI-research is a huge red flag. In general, the model is too expensive and in no way useful as is. Nor do I feel comfortable using it. This is giving me post 9/11 security theater vibes all over. Hype up fear of model capability to justify backend restrictions that violate user rights. I'm not impressed.

Said A. 的头像
Said A.4 个月前

How do you use Claude for video editing?

Dima 的头像
Dima4 个月前

Let me guess: loops?

Evgeny | Agentic Systems Builder 的头像
Evgeny | Agentic Systems Builder4 个月前

Looks like I am not testing Fable much 😭

Ofek Shaked | AI Engineer 的头像
Ofek Shaked | AI Engineer4 个月前

The real change is that verification moves upstream. You now need tighter intent specs and executable contracts before the session starts or the autonomy just executes the wrong problem faster.

Levi Figueira 的头像
Levi Figueira4 个月前

@BlnaryMlke It’s just too expensive haha Burned through my Max + 100€ of extra usage in like 1h and never finished the work… 🫠

Andrea Baccega 的头像
Andrea Baccega4 个月前

@cursor_ai bench data here is impressive for score but even more for the drop in cost per intelligence. Fable5 Med $8.27/69.8% vs Opus 4.7 Max $11.02/64.8% Fable5 Low $5.70/64.2% vs Opus 4.8 Max $7.59/63.8% The trend about the price per intelligence dropping is continuing.

Siddhant Oswal 的头像
Siddhant Oswal4 个月前

Can we use claude for video editing?

Stephen Benjamin 的头像
Stephen Benjamin4 个月前

Had it scan a project for bugs and it found a minor security thing. Instantly shut down... and now that session is toast because the security thing is in the context. I read the docs as saying I'd be downgraded to Opus?

Arian Agrawal 的头像
Arian Agrawal4 个月前

I don't know what is real or fake @trq212 anymore after the office episode

Dev K 的头像
Dev K4 个月前

"we used to verify it did the work right. now we verify it's doing the right work." that one sentence describes the entire shift in how we build.

Mr Strijker 的头像
Mr Strijker4 个月前

@mattpocockuk they basically use your grill-me skill, it is becoming more and more important as models get smarter!

Frank 的头像
Frank4 个月前

Cool, now we need a e8b and 30b mini Claude’s to run locally in conjunction with the main cloud model so it can offload minor remedial tasks while leveraging the users hardware to accelerate the user facing side

Gaurav Chande 的头像
Gaurav Chande4 个月前

Wonder where they got that idea from cc @mattpocockuk

Craig Dennis 的头像
Craig Dennis4 个月前

"If there's something you think LLMs couldn't do, give it a chance" Such good advice @trq212 !

Element Dong 的头像
Element Dong4 个月前

'verify it did the work right' vs 'verify it's doing the right work.' one is a QA problem. the other is a judgment problem.

Anders B. Eriksen 的头像
Anders B. Eriksen4 个月前

Testing it now on the same real task I gave Opus earlier today. Early impression: I’m correcting less and trusting it more. Let’s see if that holds through the final result.

Davidd Tech 的头像
Davidd Tech4 个月前

Insane stuff, congrats guys. Looking forward to testing Fable 5.

Gera92 的头像
Gera924 个月前

you stop checking if it's right, you start checking if it's even trying to do the thing you asked

Rozzabuilds 的头像
Rozzabuilds4 个月前

brb taking out a mortgage to afford a 4th Max plan

Otsukimi 的头像
Otsukimi4 个月前

“Fable can run hours at a time” Yeah for those that can causally dump $10k lol

Amin Sandolong 的头像
Amin Sandolong4 个月前

Now Claude is verifying our work, not the other way around. We’re officially just ideas guys now 😂

相关视频

Claude Tag has completely changed the way I do work for the last 4 months. Except… it's not Claude Tag. Anthropic only announced that a few hours ago, and I don't even have access yet. But I did build a version of it for myself which I've been using for months now. Here's how. 4 months ago, inspired by the success of OpenClaw, I wondered what would happen if I let Claude Code on its own computer 24x7. So I built a simple harness that allowed me to turn any Mac into an AI employee with Claude Code headless mode (-p). Today, I manage 3 such AI employees. It started with Luo Ji — my and my brother Piyush Agarwal's AI co-founder, running in our personal Slack. Luo does real work for us. We've been writing a 100% of the code for 3 products on Slack with Luo now. It manages our emails and gives us a little brief each day with things we need to take action on. And so much more. And it's not just the two of us. On the consulting team at Every 🪨, we run Claudie and for the editorial team, Andy. Same architecture, same Slack, months of real work. They help the teams with work related to project management, chief-of-staff work, data hygiene, building decks, writing first drafts, even browsing X on their own account for AI updates. So it's mindblowing to see that Anthropic landed on the exact same architecture I did. Claude Tag is an AI employee that lives in your Slack workspace and does work autonomously. Anthropic says they've been running it internally for the better part of this year — opening PRs, doing real work. And so have I. So has my whole team. The architectural decisions Anthropic baked into Claude Tag are the ones we arrived at too: - Built on Claude Code - Uses its own accounts - A separate employee per team - Slack as the interface This is the future of work, and I've been living it for months. I've shifted all of my workflows — code, PRs, even the non-technical stuff — out of Claude Code and the Claude app and into Slack. I've had entire weeks where I never opened Claude Code on my laptop. Here's a video walkthrough of how I've been using this in real life.

Nityesh

36,790 次观看 • 3 个月前