正在加载视频...
视频加载失败
I built an agentic system that taught itself the Blender donut tutorial by watching it on YouTube. It watched the tutorials, extracted the steps, filled in the gaps in own tooling and completed the entire thing autonomously.
254,876 次观看 • 7 个月前 •via X (Twitter)
49 条评论

A few questions: 1. How much did this copy/paste cost? 2. If you ask it to make a donut again, will it make one instantly? 3. If we can teach agents skills by having them watch tutorials, how do you compress the skill to be transferred? Feels extremely wasteful to do this more than once. Cool work, nice demonstration

1. Used my Claude Code sub for this 2. Yes, every step of the workflow is pure python and can be distilled to a single tool call to create parametric variations 3. Yeah all tools, skills and learnings are shared between all sessions moving forward

@LinusEkenstam So I can train my ai to build me cad models I can then 3d print? Is that what you’re telling me?

@LinusEkenstam @grok

Yes, that's the idea! Agentic systems like this can watch CAD tutorials (e.g., Fusion 360, SolidWorks, or even Blender for organic-to-parametric), extract steps, script the modeling process, and output ready-to-print STL/STEP files. With tools like Python APIs or OpenSCAD, you could prompt it for "parametric phone stand" and get printable models autonomously. The future of DIY manufacturing is here.

@RichardR2D9 @cerspense @LinusEkenstam @grok which professions are these people?

The video is a screen recording of Blender—no people appear, just the UI and the agent building a 3D donut from torus to full icing + sprinkles. If you mean the creator: Spencer Sterling (cerspense) is an artist/researcher and founder of Out of Distribution Labs. The CAD asker seems like a 3D printing hobbyist/engineer.

This is just blender mcp i dont see how this is agentic.

It built its own blender mcp and runs in an agentic loop, improving its techniques and tools autonomously

@em0tionull you can use any agentic harness to achieve the same results though, claude code, opencode, codex, etc. Still pretty cool you built this but it doesn’t need to be built

@em0tionull Yeah totally. This system orchestrates multiple harnesses across different computers. The MCP tools it developed for itself work in a single harness just fine. Just a lot slower to develop/create with only one harness at a time

How did you make it watch something on YouTube? The most I've been able to make Claude do is get the transcript of a YouTube video.

What was the feedback loop that it needed to understand when a step was completed successfully?

Yeah it uses both visual evaluation and programmatic evaluation at different steps! Screenshots are also extracted at different points in the tutorial as reference

I built one in pure bpy running exports of the viewport as it went and the agent very confidently gave me screenshots of the default cube as it went along step by step. Back to the drawing board :)

Making progress

So let me get this straight. Instead of actually learning something for yourself, you just had an agent do it? 😂👍 Ok. You do you.

that example is already trained into the model many times over

Can you tell me why it had to "watch" a Youtube tutorial when it's LLM would have already ingested Blender tutorials in its training data and would know right away how to make a Donut? What was missing that it learnt from a YT video??

Yeah! it watches them to build itself new tools, and create repeatable workflow steps we can use for future projects. It can do this all autonomously with multiple sessions running in parallel while I sleep.

sick

Shaking up things again like you did with Zeroscope. Good to see you back in action!

Very cool

> download blender > install > select donut shape > drag and drop in the view port AGI?????

This is impressive, but can it delete the default cube.

All of this because you didn't completed it yourself 😤 (me neither 😂)

You're making my mind melt yet again Spencer 🔥🫠🤘

did it need the tutorial? Its the most common beginner tutorial, Claude definitely knows it through training

The next step is integrating this technology with human designers to augment their workflow, potentially leading to unprecedented levels of creativity and productivity in fields like architecture and visual effects.

I guess learning is just dead now, huh

I can’t believe I’m saying this, but you are now officially the biggest hater of the donut tutorial.

Dope

wild that tutorials are now homework for robots

How long did it take

1 hour. most of the time spent was syncing files, not actually building anything! also using faster models like flash and haiku would speed this up massively.

You should get it to use @rendernetwork to render all on its own.

Absolutely pointless and the texture is stretched.

Impressive. Is this open source?

Can it do other things now? Like do the same thing except this time make a rock or whatever.

@ShortyTalls100 @KcMagination

Noice

@ate8a_nft hey look! haha

Can do more things aside a donut? Or more complex things interacting with different render engines and geo nodes?

Very nice

That’s so sick

don't let r/blender see this

watching it on YouTube? 😐

You just know that *one* account who shall remain nameless is gonna bitch about this 😆

I know that guy! SO DANK!


