Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

I gave Fable 5 one job: write custom WebGPU kernels for Gemma 4 inference. It climbed to 84 tok/s, then hit a wall, insisting further optimization was impossible. Hours later, Anthropic rolled back invisible LLM development safeguards, and it hit 255 tok/s. The next day, access to Fable 5...

1,168,280 Aufrufe • vor 3 Monaten •via X (Twitter)

34 Kommentare

Profilbild von Neelakandan NC
Neelakandan NCvor 3 Monaten

Soon China is going to build a model like fable and it is going to be open - which will make everyone lean on China than us

Profilbild von Gordon Olson
Gordon Olsonvor 3 Monaten

I have been a fan of webgpu for sometime. Here is one for you. Browser-native persistent AI agent runtime with local memory, context reconstruction, compiled WebGPU inference, and a custom WebGPU Kernel Lab.

Profilbild von Stephen
Stephenvor 3 Monaten

This is fake — the invisible safeguards being removed means it went over to Opus instead of sandbagging

Profilbild von Sina Shahandeh
Sina Shahandehvor 3 Monaten

Their fear mongering strategy backfired. When they compare AI with nuclear bombs, now they have deprived all of advanced AI. If the government start to control AI like this, it will slow down all progress, not unlike the middle ages, where thought was controlled by the church.

Profilbild von Nick Dobos
Nick Dobosvor 3 Monaten

How do you know it was an invisible classifier? And not simply the model having a break through discovery? Seems viable to have natural plateaus and stalls in progress sometimes.

Profilbild von Foreman Panda
Foreman Pandavor 3 Monaten

But can this only be done by Fable 5? Any comparison?

Profilbild von Lee Penkman
Lee Penkmanvor 3 Monaten

i did this but with my stock trading bot lmao. pretty selfish of me but now its getting near 3x returns/mo so thats great. also started on a minecraft clone then ran out of usage

Profilbild von Christopher
Christophervor 3 Monaten

This is funny because the safeguards themselves were never rolled back

Profilbild von Migel Tissera
Migel Tisseravor 3 Monaten

Curios to know, did you try this with GPT-5.5? What happened?

Profilbild von alp
alpvor 3 Monaten

Anthropic did not roll back invisible safeguards. It just rolled out the invisible part.

Profilbild von Asher Crowe 🪺
Asher Crowe 🪺vor 3 Monaten

the kernel optimizing itself overnight while you slept is mildly terrifying and i'm here for it

Profilbild von Kirk Patrick Miller
Kirk Patrick Millervor 3 Monaten

All of this is about stopping the small and protecting monopolies. None of this is about safety. •

Profilbild von Lon Lundgren
Lon Lundgrenvor 3 Monaten

Now just plug the same prompt into Opus and you can watch it get the same result (which is all Fable was doing without the safeguards).

Profilbild von Eyal Toledano
Eyal Toledanovor 3 Monaten

Doin Gemma 4 today too Been doin agentic kernel optimization since mid last year. Those last winning tests are Fable 5 right before it was suspended. It was on a tear 😭

Profilbild von AI Mastery Guide
AI Mastery Guidevor 3 Monaten

84 to 255 tok/s the moment safeguards rolled back is a wild data point. That gap between what the model can do and what it's allowed to do is bigger than most people realized.

Profilbild von Akash
Akashvor 3 Monaten

what if you give same tasks to gpt5.5 (xhigh)?

Profilbild von Sacrificial Pancakes
Sacrificial Pancakesvor 3 Monaten

I was like 75% solving a pet problem I’ve been chipping away at for years… “this model is unavailable” 😭

Profilbild von cryptovatooor - (Techno, Optimist)
cryptovatooor - (Techno, Optimist)vor 3 Monaten

fascinating

Profilbild von Brandon
Brandonvor 3 Monaten

Woah

Profilbild von luka
lukavor 3 Monaten

okay, I'm interested to know how often it was making edits or being looped during this period. For instance between minute 100 and 180

Profilbild von πανοπλίαν τοῦ Θεοῦ
πανοπλίαν τοῦ Θεοῦvor 3 Monaten

I’ve been working on a similar project, and just ported it to a standalone was able to get qwen3.6-35b-3ab on 8gb ram, qwen3.6-27b on a 4090, and qwen3.5-27b on an rx 6700 xt, also worked with a 1080ti

Profilbild von John D. Pope  🦒
John D. Pope 🦒vor 3 Monaten

@peteskomoroch Can you share your code? Help liberate the models? Maybe opus can use it elsewhere

Profilbild von Sebastian Buzdugan
Sebastian Buzduganvor 3 Monaten

255 tok/s after warmup means little because webgpu shader cache misses dominate first token

Profilbild von Jeffrey Castellano
Jeffrey Castellanovor 3 Monaten

I was doing exactly this too, I did the Gemma 4 kernels a month ago but in one day I got Gemma 12B and diffusion working with WebGPU in my runtime, diffusion got the plug pulled in the final hour last night. They pulled the plug and 4.8 Opus told me it was impossible.

Profilbild von CryptoCow
CryptoCowvor 3 Monaten

This is heckin insane if you were able to actually do this!!? What about for like qwen?

Profilbild von qalqi.com
qalqi.comvor 3 Monaten

-- Using Self Improving Agent framwork paired with unsloth.. wouldn't it be possible to spawn and train a slm for this?

Profilbild von Jonathan Leaders
Jonathan Leadersvor 3 Monaten

Technically it wasn't banned globally. It was banned for non-Americans and then they withdrew it from everyone.

Profilbild von steve
stevevor 3 Monaten

did u notice in the previous period that 4.8 was actually getting called?

Profilbild von ani4ani
ani4anivor 3 Monaten

Did you ask it write a report of what it tried first and then what changed after it started increasing again so you know what was gated

Profilbild von Matt Newell
Matt Newellvor 3 Monaten

Anthropic only changed what happened when you hit safeguards (previously, "shadowban"; now, Opus). You just placed the kink in the straight line in such a place to make it look like this was related to the change.

Profilbild von Sergio Suave
Sergio Suavevor 3 Monaten

So you were basically using Opus in the background until they rolled back the guardrails...

Profilbild von Austin
Austinvor 3 Monaten

@grok, what are all the optimizations discovered in the video

Profilbild von Jean-Paul Tres
Jean-Paul Tresvor 3 Monaten

Repo link? 🧐

Profilbild von JulianSaks
JulianSaksvor 3 Monaten

🥲

Ähnliche Videos

fable 5.1 vs fable 5 vs opus 5 – three lord of the rings landmarks, built in 3d from one image the setup: one reference image per scene, one html file per build, everything procedural – no meshes, no textures, no image files, nothing past Three.js from a cdn. each model reads the picture, writes its own prompt from it, then builds to that prompt in the same turn. three named camera shots per scene on keys 1/2/3, so it can be screen-recorded. run through OpenRouter tasks: 1. bag end – hobbiton from two frames, outside and in. the round green door has to open onto the room you are standing in 2. barad-dûr – the tower and orodruin from one film still. the eye has to move and track the camera, the volcano erupts on a cycle, the clouds never stop 3. rivendell – jerry vanderstelt's painting. sun shafts that shimmer, water that falls without a break, trees that sway on a gust models: Anthropic fable 5.1, fable 5, opus 5 total cost, three builds #1 fable 5 – $14.97 #2 opus 5 – $18.53 #3 fable 5.1 – $22.38 wall clock, three builds #1 fable 5 – 38m #2 fable 5.1 – 92m #3 opus 5 – 122m output tokens #1 fable 5 – 298,592 #2 fable 5.1 – 439,435 #3 opus 5 – 724,418 lines of code shipped #1 fable 5 – 2,885 #2 fable 5.1 – 4,021 #3 opus 5 – 5,161 biggest single build, lines #1 opus 5, bag end – 2,410 #2 fable 5.1, barad-dûr – 1,375 #3 fable 5, bag end – 1,319 observations: • fable 5.1 is the only model that furnished the bag end interior – a live fire, panelling, books on the floor, leaded diamond windows, against fable 5's flat color and opus's dark tunnel. the round door outside opens onto that room, the hard part of the brief • what it costs is thinking room. the 128k output ceiling is a thinking budget in disguise: fable 5.1 burned 102,116 of it on reasoning and hit the wall mid-file. opus spent 109,241 and hit the same wall. fable 5 spent 61,240 and finished bag end in one call – the only one that did • fable 5.1's first pass is not the finished thing. its barad-dûr came back with three defects you only catch by looking at it – nothing a read of the code would have flagged • it is the best of the three at being corrected. handed a plain list of what was wrong, it returned 32 targeted patches over two rounds, every one applied first try, and it worked out one of the causes itself instead of guessing at constants conclusion: nine scenes, 12,067 lines and 1.46m output tokens for $55.88 all in – and the cheapest model was also the fastest, by 3.2x! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

18,509 Aufrufe • vor 26 Tagen

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 Aufrufe • vor 3 Monaten