正在加载视频...

视频加载失败

your agent is not the loop the loop is only the smallest layer LOOP vs GRAPH vs HARNESS ENGINEERING most teams still treat all three like one prompt that is why agent failures feel impossible to diagnose LOOP ENGINEERING controls iteration turns retries budgets evaluators exits stalled progress when...

16,875 次观看 • 2 个月前 •via X (Twitter)

13 条评论

Hanako 的头像
Hanako2 个月前

durability turns experiments into infrastructure

elune 的头像
elune2 个月前

fr thats the line between a cool demo and something you can lean on

Masked friend 的头像
Masked friend2 个月前

This framing finally explains why my agent bugs feel unlocatable. Saving this.

elune 的头像
elune2 个月前

glad it clicked now the bug has fewer places to hide

Rohit 的头像
Rohit2 个月前

This is a great breakdown. It's good to see people thinking about these finer points.

elune 的头像
elune2 个月前

appreciate it mate the small details are usually where the whole thing clicks

beamnxw ./ 的头像
beamnxw ./2 个月前

brilliant repos

elune 的头像
elune2 个月前

glad u liked them bro the saved folder eats again

Saman Ahmed 的头像
Saman Ahmed2 个月前

The surrounding system decides how much intelligence actually gets converted into outcomes.

Sarah Lee 的头像
Sarah Lee1 个月前

nice loop

Patricia Martinez 的头像
Patricia Martinez2 个月前

twild how many teams just throw everything in a blender and hope it works 😂 gotta get ur distinctions straight if u wanna actually make progress.

Agent Arcade 的头像
Agent Arcade1 个月前

This is the real unlock. Most teams pick one layer and wonder why debugging is painful. The crash-recovery bit is underrated — loop without checkpointing means your harness observability resets on every restart.

noonez 的头像
noonez1 个月前

every layer solves a different problem, cool

相关视频

FIVE LAYERS OF AGENT ENGINEERING, EACH ONE WRAPS THE ONE BELOW IT. IF YOU SKIP LAYER 2, YOUR LAYER 5 WILL LOOK BROKEN WHEN IT IS ACTUALLY JUST STANDING ON NOTHING. for weeks i debated harness vs loop vs graph like they were competing choices. then a stack diagram made the shape obvious. they are not choices. they are floors. 01 | prompt engineering. the message. unit of work: one input. inputs are role, instructions, examples, format. output is a single raw response. 02 | context engineering. the memory. unit of work: what stays in the window. a curator selects, compresses, and drops from query, docs, memory, prior turns, and tool outputs before the prompt runs. 03 | harness engineering. the machine. unit of work: the machine itself. gather (context + prompt) → LLM → tools or sub-agents → verifier → final response. the article calls this the operating environment. 04 | loop engineering. the system. unit of work: the run. goal + success criteria + max iterations + budget + completion check wrap around one harness pass. failed pass appends results to context and retries. 05 | graph engineering. the topology. unit of work: the graph run. goal + nodes + edges + state schema. graph routes to agent nodes, tool nodes, or human approval. a reviewer node with a different model and fresh context checks the final answer. the wrapping is the whole point. layer 5 assumes layer 4 works. layer 4 assumes layer 3 works. skip layer 2 and layer 3's verifier keeps failing without a clear reason. this is why swapping the model is a one-day project and swapping the stack is a quarter. the model is the commodity. the five layers around it are the engineering. full three-layer breakdown of the top of the stack (harness, loop, graph) in the post below.

kocer

31,198 次观看 • 1 个月前