Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

This is what the future of coding looks like. Multiple AI models working in the same Codex session, with Jev routing them to tackle different parts of the task. Here, I had it build a full-stack demo app: > Opus 5.5 planned it. > GPT-6 Astra built the backend....

15,765 Aufrufe • vor 1 Tag •via X (Twitter)

13 Kommentare

Profilbild von Manu Agrawal
Manu Agrawalvor 1 Tag

Very interesting usecase. Loved your solution 🙌

Profilbild von Viber · fireply.ai
Viber · fireply.aivor 1 Tag

how many tokens did opus burn just planning a todo app lol

Profilbild von Casey Lee Race
Casey Lee Racevor 1 Tag

How do you handle the distributed race conditions even with Jev?

Profilbild von Josh Hoeg | Vibe Coding & AI
Josh Hoeg | Vibe Coding & AIvor 1 Tag

never touching the model picker again sounds great ngl

Profilbild von 🦋 novababe
🦋 novababevor 1 Tag

Looks like I finally found a team that actually listens to my “let's refactor” suggestions. ✨

Profilbild von Blum
Blumvor 1 Tag

Matching each task to the model that handles it best sounds like a damn effective approach

Profilbild von Francisco Moreno Diaz
Francisco Moreno Diazvor 1 Tag

Buen ejemplo de enrutar modelos por pedazo de la tarea. En contaduría sirve igual: uno arma la conciliación, otro revisa excepciones. El control sigue en la bitácora y en quien firma el corte.

Profilbild von Peter Orban
Peter Orbanvor 1 Tag

Splitting one Codex session across planners and builders is the part I want more of. Opus on the plan, Astra on the backend, Kimi on the rest only works if the router actually reads the whole task instead of the first line of the prompt.

Profilbild von Yann Kronberg
Yann Kronbergvor 1 Tag

That 40% cheaper number is what I'd keep watching over a longer run, because once five models are taking turns you can't assume the cache follows them, and some of those savings can disappear just to keep each model up to speed.

Profilbild von Hashim Syed
Hashim Syedvor 1 Tag

this is exactly what we built it for. one session, the right model on each part, and you never touch the picker. thanks for putting it through a real build

Profilbild von 安叫兽|Bird🕊️ 🔶 BNB
安叫兽|Bird🕊️ 🔶 BNBvor 1 Tag

以后写代码像组建临时项目组了娱乐主管

Profilbild von Gregor
Gregorvor 1 Tag

Opus planned it and GPT built it. The builder never had the planner's full context.

Profilbild von shaped
shapedvor 1 Tag

this is pretty cool, and it kinda means the overall leaderboard matters less now but what we really needa focus on is which model's best at each piece of the job

Ähnliche Videos

Chinese AI models are wiping billions off Big Tech right now. Google just lost $200 billion in a single day, and the model it needed to fight back still isn't ready. Gemini 3.5 Pro, Google's most powerful model, is months behind schedule. Alphabet stock dropped 4.4% that same day. The Deepseek moment is happening again, and the new model is FAR bigger. On the same day Google's delay leaked, a Beijing lab called Moonshot released Kimi K3. It is the largest open model ever built, with 2.8 trillion parameters. It took the number one spot on the Frontend Code Arena, a live coding leaderboard, passing Anthropic's best model. And Moonshot is giving it away for free on July 27. The genius part: Anyone with enough computers can download it and run a frontier level AI without paying a cent to a US company. A single task on Kimi K3 costs about 94 cents. The same work on some American models costs nearly double. So why would a company keep paying premium prices for a model it can now get for free? The entire US AI business is built on selling access to models that cost billions to train. If a free Chinese version does most of the same work, that pricing power starts to crack. And Kimi is close to the best. On one closely watched intelligence ranking it scored 57, just behind the top American models GPT-5.6 Sol and Fable 5, and ahead of Claude Opus 4.8. Bank of America told clients that Kimi proves Chinese labs can keep making big leaps even with limited chips. And the founder of Moonshot, Yang Zhilin, learned to build AI as a researcher INSIDE Google. Google literally wrote the 2017 paper that made all of these models possible. Now the people who studied its work are using it to destroy Google, and handing it out for free. What happens next: Kimi K3's weights go public on July 27. Google reports earnings on July 22, and everyone will be asking the same question about Gemini. If free models keep topping the charts, every valuation built on paid AI access has to be rewritten. What do you think?

Ricardo

47,790 Aufrufe • vor 2 Monaten