Video yükleniyor...
Video Yüklenemedi
We made Fable 5.1 compete against Entelligence Router on the same coding task. Both had the same harness and infra, and had to build a GeoGuessr-style game from scratch, implement the interactions, and verify everything in the browser. Fable 5.1: 21 min, $22.16 Entelligence Router: 21 min, $9 Fable... show more
16,627 görüntüleme • 11 gün önce •via X (Twitter)
16 Yorum

why do I feel that model router performed better here 👀

And that's how you do advertising! Really interesting results!

curious how the browser verification works. is fable 5.1 using a vision model to check the rendered output or do you step in manually. the verify loop is where my agents eat the most time

$22 versus $4 is a massive difference for similar coding work

一分钱一分货 😁

Great benchmark. Same time, nearly 60% lower cost is a serious efficiency win.

Head-to-heads like this are gold 👏

Same result in 21 mins, but Entelligence Router doing it at 59% lower cost? That's incredible efficiency without compromising quality!

Same harness, same infra, clean benchmark. — corp

Impressive same result in same time but 59% cheaper

Real-world cost comparisons like this make model selection much clearer

Great benchmark healthy competition drives better, more efficient AI solutions.

Great benchmark showing how cost efficiency matters alongside coding quality```

same time, 59% cheaper is the part worth digging into. was the router mostly picking smaller models for the boilerplate steps, or is the savings coming from fewer retries?

Great comparison showcasing the evolution of AI coding capabilities! 🚀 Benchmarking real-world tasks helps push innovation and improve developer experiences. Would love to explore collaboration opportunities and support the growth of this exciting AI ecosystem. 🤝

That's a really rigorous test, and I appreciate the detailed breakdown. Fable 5.1 shows significant benchmark improvements for coding. We actually went deeper on those metrics here:
