Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

anthropic engineers just built a full app from scratch with a loop of agents, start to finish in 40 minutes it's the claude code team. they used three agents: one plans, one builds, one checks the work. they keep cycling until the app actually runs the model matters less...

15,267 Aufrufe • vor 4 Tagen •via X (Twitter)

18 Kommentare

Profilbild von Blum
Blumvor 3 Tagen

Great workshop, watched the whole thing and it was worth every minute

Profilbild von Dean W. Perkins
Dean W. Perkinsvor 3 Tagen

thanks dude

Profilbild von Martin Lepka
Martin Lepkavor 4 Tagen

I like the plan / build / check split. Plan agent is the one I'd watch tho.. a bad brief just gets built faster now :)

Profilbild von ares. 🎧
ares. 🎧vor 4 Tagen

lo que me parece más útil es que el que revisa no sea el mismo que escribe, si no se da el visto bueno a sí mismo

Profilbild von Chesny
Chesnyvor 4 Tagen

lo interesante es que siguen ciclando hasta que arranca, no se conforman con la primera versión

Profilbild von Dean W. Perkins
Dean W. Perkinsvor 4 Tagen

that's the part that makes it work, the first version is never the one that runs

Profilbild von Leo
Leovor 4 Tagen

saving this for later, nice content dude

Profilbild von Dean W. Perkins
Dean W. Perkinsvor 4 Tagen

appreciate it, worth the 40 minutes

Profilbild von Jatin Garg
Jatin Gargvor 4 Tagen

the 40-minute end-to-end is interesting. how did they handle the agent failures - retry with same prompt or route to a different agent?

Profilbild von roman
romanvor 4 Tagen

lo de que el modelo importa menos que el bucle es la frase que más gente debería leer hoy

Profilbild von Marco
Marcovor 4 Tagen

me lo guardo, justo estoy montando algo parecido y lo tengo todo en un único agente

Profilbild von Dean W. Perkins
Dean W. Perkinsvor 4 Tagen

I’d love to see it when you’re finished with it

Profilbild von marcus
marcusvor 4 Tagen

que lo enseñe el propio equipo que construye la herramienta le quita toda la duda

Profilbild von Dean W. Perkins
Dean W. Perkinsvor 4 Tagen

hard to argue with the team that builds the thing

Profilbild von catman
catmanvor 4 Tagen

The loop had distinct jobs: one agent planned, one built, and one checked whether the app actually ran. That validation step is what closes the gap between generated code and a working app.

Profilbild von crsn
crsnvor 4 Tagen

the checker agent is doing the heavy lifting here. planner and builder are easy, getting something that actually says "no, redo it" and is right about it is the hard part

Profilbild von astro
astrovor 4 Tagen

@aiscwork barbaridad de vídeo, todo el mundo debería verlo

Profilbild von Hunit
Hunitvor 3 Tagen

A planner, a builder, and a checker looping until the app actually runs is a neat pattern, and it fits the idea that the harness around the model matters as much as the model itself. You and @mlkbfrey are the two accounts I enjoy following most.

Ähnliche Videos