正在加载视频...

视频加载失败

Boris Cherny, Head of Claude Code, Anthropic: "you can define a stop hook that's like, if the tests don't pass, keep going. essentially it's like you can just make the model keep going until the thing is done." that one config is the post's loop #1, straight from the...

30,025 次观看 • 3 个月前 •via X (Twitter)

7 条评论

Zhen 的头像
Zhen3 个月前

The pattern underneath all four loops is the same: make output verifiable while it is still being produced, not after. A stop hook forces tests green before the turn ends; a fresh-context critic catches what the builder talked itself past. The interesting transfer is that the same readable-while-it-happens bar applies beyond code, to any live output a human has to follow in real time.

安叫兽|Bird🕊️ 🔶 BNB 的头像
安叫兽|Bird🕊️ 🔶 BNB3 个月前

这招有点狠,别把额度也循环没了

zyxron 的头像
zyxron3 个月前

Oh, this guy definitely knows what he's talking about

Harlan A Nelson 的头像
Harlan A Nelson3 个月前

Interesting that is seems the Anthropic engineers have never heard of bash bang. That’s the difference between someone who started programming when the main tool available was the command line vs people who started with a gui.

Sigo Egwey 的头像
Sigo Egwey3 个月前

Hooks are so powerful in claude code. highly underrated

Lakshmi 的头像
Lakshmi3 个月前

An agent that fixes your GitHub repo and opens a PR LangGraph + self-improving loop. Live & open source. ▶ ⭐

也无风雨也雾晴 的头像
也无风雨也雾晴3 个月前

That stop hook is powerful, but I’d still want the loop drawn somewhere visible. For agent workflows, the useful map is: trigger -> model action -> test/check -> human review boundary. Small related tool I use for these maps:

相关视频

HOW TO USE AI LOOPS TO RUN YOUR BUSINESS 24/7 A lot has been written about loop engineering for building products. Almost nothing about using loops to run the business itself. That's the bigger idea. A loop is when you give an agent a goal, a way to check its own work, and permission to keep trying until it hits that goal. Build. Verify. Repeat. Stop when the condition is met. Here's what it looks like in practice: 1/SEO loop You're position 30 for a term you want. The loop runs once a month, makes changes, checks where you rank, and keeps pushing until you're on page one. This is running in production right now on Inbox Zero. 2/Ads loop You're spending $100 a day and losing money. The loop tests creative, checks profitability, kills what fails, and keeps going until the account is in the black. 3/Eval loop Your AI feature is only 88% accurate. The loop keeps adjusting the prompt and swapping the model until it passes 90%. 4/LLM visibility loop People search in ChatGPT now, not just Google. Same loop, new scoreboard. Are we the answer or not? The whole thing hinges on one thing: a metric that comes back black and white. Where do I rank? Did it hit profitability? Did the evals pass? Give an agent that scoreboard and it runs for months. Loops used to run for 30 minutes. These run for a year. Take a step, sleep, wake up next month, take another one. You're basically hiring an agency that never sleeps, gets paid in tokens instead of invoices, and undoes its own mistakes when the number goes down. Full episode on The Startup Ideas Podcast (SIP) 🧃 watch

GREG ISENBERG

83,349 次观看 • 2 个月前

this is rare f*cking gold an internal AI engineering document leaked. it's saving solo devs $300,000 a year that number is just the three hires you never make: the one who writes the spec, the one who reviews the output, the one who runs the queue prompt-driven is out. loop-driven is in. that shift changes everything Generate → Evaluate → Remember → Schedule → Optimize → Recurse six layers, one loop. it improves itself. no human in the middle > generation - writes its own brief, then builds against it > evaluation - a second layer grades the work and hands it back > memory - keeps what worked, so the same mistake costs you once > scheduling - picks the next job. the queue runs itself > optimization - rewrites its own instructions from what shipped > recursion - pull one layer out and the whole thing degrades that last one is the tell. six layers is the floor, not the wishlist here's the career part nobody says out loud every team already has someone typing prompts all day. nobody has the person who can build the loop that replaces that job bring this into your company and you stop being the one doing the work. you become the one who designs how the work runs that's the jump from operator to architect, and it's the promotion that's actually open right now AI stops predicting the next token. it starts running its own loop the people who wire this up over a weekend spend next year reviewing output. everyone else keeps typing prompts

Annatar.md

26,968 次观看 • 21 天前

THIS GUY CONNECTED HIS AI AGENTS TO HIS OBSIDIAN AND BUILT A BRAIN THAT LEARNS ON ITS OWN. HERE'S HOW TO BUILD IT Obsidian is just markdown files sitting in a folder. That turns out to be the perfect memory for an AI agent, because an agent can read and write those files directly. He wired his agents into the vault so they pull context from it, do the work, and write what they learned back. The notes aren't the point. The loop is, and it gets sharper every cycle How to build it: 1. Point an agent at your vault. The fastest way, no plugins, no API keys: open a terminal and run npx obsidian-mcp /path/to/your/vault. That exposes your Obsidian folder to Claude as a tool it can read, search, and write to. Add it to your Claude Code or Cowork config and restart 2. Confirm it can see the brain. Ask it: "list the notes in my vault and summarize what's in them." If it reads them back, the connection is live. Now it starts every task with everything the vault already holds instead of from zero 3. Give each agent one job and a write-back rule. Tell it: "research this, then save what you found as a new note in /brain with links to related notes." One agent researches, one summarizes, one plans. Each writes its output back into the vault 4. Close the loop. Add one line to every agent's instructions: "read /brain before starting, write your result back when done." Now each task leaves the vault richer, and the next run reads that before it works. It compounds instead of resetting 5. You only steer. Review what the brain produces, point it at the next thing. The agents handle the reading, writing, and connecting The edge isn't better notes. It's a brain that feeds itself, so the work gets sharper every cycle instead of starting over Bookmark this

Yarchi

58,643 次观看 • 4 个月前