Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Anthropic CEO after realizing a Chinese dev explained loop vs graph engineering better than anyone at the company

1,422,754 Aufrufe • vor 12 Tagen •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

What feels better than this after a long day at the gym?
0:15

Sensitive content

What feels better than this after a long day at the gym?

Cherri Findom

54,222 Aufrufe • vor 1 Monat

anthropic will sell you opus 5 at $200/mo. openai will sell you gpt-5.6 at $200/mo. neither will tell you the fix that drops your bill to $20 was posted free on langchain's blog on july 18 peter steinberger posted one line asking if we'd moved from loops to graphs yet. 24 hours later there was a manifesto. a week later every ai account had a $497 graph engineering course. all of them wrong about the same thing the sentence that ends the argument, buried in a langchain post nobody quoted: loop engineering isn't an alternative to graphs, so much as a simple version of them the machine, five layers, each wraps the one below: L1 the ask · 23% of errors (anthropic red team, q4 2024) -> "just add more instructions" burns tokens with zero accuracy gain -> real fix: examples, output schema, constraints as positives L2 the context · where 90% of you actually die -> 140,500 tokens where 18,000 would work, 8x the price for the worse answer -> real fix: retrieve, rank, compact, clear dead tool outputs L3 the harness · 31% of "model bugs" are harness bugs (openai safety eval, 2024) -> unbounded file perms = avg $23,400 incident. sandboxed = $0 (stripe internal) -> no timeout = $847 median in api fees before you notice -> real fix: explicit scopes, timeouts, human-required gates L4 the loop · "it stopped" is a loop exit problem -> the verifier said "looks good" to garbage. again -> real fix: machine-checkable exit test, turn cap, rubric L5 the graph · only 12% of teams use graphs in prod (stanford hai, n=2,841) -> 58% of graph failures are wrong-agent selection, not model -> teams abandon graphs saying "harder to debug than a loop." that's a harness problem -> real fix: name every node's specialty, delete decoration fix down, not up. a symptom at layer 4 usually originates at layer 2. a bigger model on a broken harness is a smarter employee locked in the same empty room drop your $200/mo ai sub to $20, check the article below

starmex

145,325 Aufrufe • vor 18 Tagen