Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

This is what the future of coding looks like. Multiple AI models working in the same Codex session, with Jev routing them to tackle different parts of the task. Here, I had it build a full-stack demo app: > Opus 5.5 planned it. > GPT-6 Astra built the backend....

15,765 görüntüleme • 1 gün önce •via X (Twitter)

13 Yorum

Manu Agrawal profil fotoğrafı
Manu Agrawal1 gün önce

Very interesting usecase. Loved your solution 🙌

Viber · fireply.ai profil fotoğrafı
Viber · fireply.ai1 gün önce

how many tokens did opus burn just planning a todo app lol

Casey Lee Race profil fotoğrafı
Casey Lee Race1 gün önce

How do you handle the distributed race conditions even with Jev?

Josh Hoeg | Vibe Coding & AI profil fotoğrafı
Josh Hoeg | Vibe Coding & AI1 gün önce

never touching the model picker again sounds great ngl

🦋 novababe profil fotoğrafı
🦋 novababe1 gün önce

Looks like I finally found a team that actually listens to my “let's refactor” suggestions. ✨

Blum profil fotoğrafı
Blum1 gün önce

Matching each task to the model that handles it best sounds like a damn effective approach

Francisco Moreno Diaz profil fotoğrafı
Francisco Moreno Diaz1 gün önce

Buen ejemplo de enrutar modelos por pedazo de la tarea. En contaduría sirve igual: uno arma la conciliación, otro revisa excepciones. El control sigue en la bitácora y en quien firma el corte.

Peter Orban profil fotoğrafı
Peter Orban1 gün önce

Splitting one Codex session across planners and builders is the part I want more of. Opus on the plan, Astra on the backend, Kimi on the rest only works if the router actually reads the whole task instead of the first line of the prompt.

Yann Kronberg profil fotoğrafı
Yann Kronberg1 gün önce

That 40% cheaper number is what I'd keep watching over a longer run, because once five models are taking turns you can't assume the cache follows them, and some of those savings can disappear just to keep each model up to speed.

Hashim Syed profil fotoğrafı
Hashim Syed1 gün önce

this is exactly what we built it for. one session, the right model on each part, and you never touch the picker. thanks for putting it through a real build

安叫兽|Bird🕊️ 🔶 BNB profil fotoğrafı
安叫兽|Bird🕊️ 🔶 BNB1 gün önce

以后写代码像组建临时项目组了娱乐主管

Gregor profil fotoğrafı
Gregor1 gün önce

Opus planned it and GPT built it. The builder never had the planner's full context.

shaped profil fotoğrafı
shaped1 gün önce

this is pretty cool, and it kinda means the overall leaderboard matters less now but what we really needa focus on is which model's best at each piece of the job

Benzer Videolar

Chinese AI models are wiping billions off Big Tech right now. Google just lost $200 billion in a single day, and the model it needed to fight back still isn't ready. Gemini 3.5 Pro, Google's most powerful model, is months behind schedule. Alphabet stock dropped 4.4% that same day. The Deepseek moment is happening again, and the new model is FAR bigger. On the same day Google's delay leaked, a Beijing lab called Moonshot released Kimi K3. It is the largest open model ever built, with 2.8 trillion parameters. It took the number one spot on the Frontend Code Arena, a live coding leaderboard, passing Anthropic's best model. And Moonshot is giving it away for free on July 27. The genius part: Anyone with enough computers can download it and run a frontier level AI without paying a cent to a US company. A single task on Kimi K3 costs about 94 cents. The same work on some American models costs nearly double. So why would a company keep paying premium prices for a model it can now get for free? The entire US AI business is built on selling access to models that cost billions to train. If a free Chinese version does most of the same work, that pricing power starts to crack. And Kimi is close to the best. On one closely watched intelligence ranking it scored 57, just behind the top American models GPT-5.6 Sol and Fable 5, and ahead of Claude Opus 4.8. Bank of America told clients that Kimi proves Chinese labs can keep making big leaps even with limited chips. And the founder of Moonshot, Yang Zhilin, learned to build AI as a researcher INSIDE Google. Google literally wrote the 2017 paper that made all of these models possible. Now the people who studied its work are using it to destroy Google, and handing it out for free. What happens next: Kimi K3's weights go public on July 27. Google reports earnings on July 22, and everyone will be asking the same question about Gemini. If free models keep topping the charts, every valuation built on paid AI access has to be rewritten. What do you think?

Ricardo

47,790 görüntüleme • 2 ay önce