正在加载视频...

视频加载失败

Vibecoding ➡️ Vibesculpting (in 3D) I generated a 3D model of a Spitfire-looking plane in Blender, using Claude 3.7 + Blender MCP (link below 👇), piece by piece... using natural language only. Seeing this alongside Google's new image edition demo, I think we're ahead of a true seismic shift...

95,572 次观看 • 1 年前 •via X (Twitter)

12 条评论

Emm 🔜 GDC 2025 的头像
Emm 🔜 GDC 20251 年前

Start to finish..

Emm 🔜 GDC 2025 的头像
Emm 🔜 GDC 20251 年前

I used by @sidahuj - try it out below.

Digital Currency 的头像
Digital Currency2 年前

From 3D modeling to VR/AR development, our MSc in Metaverse program equips you with the technical skills to excel in the rapidly evolving digital world. Don't miss out—enroll today! #UNIC #MScMetaverse

George Crudo 的头像
George Crudo1 年前

why is "Claude" manually clicking drop down menus and manipulating object with the gizmo when the rest of the time its sending scripting commands to Blender? 🤔

Emm 🔜 GDC 2025 的头像
Emm 🔜 GDC 20251 年前

Glad you're paying attention, George! 😂 Bc Claude doesn't know - yet - how to properly position four propeller blades together in Blender. Same thing with the cockpit, the cowling and the vertical tail - clicking was faster than prompting. I also wanted it to be perfect 48 hours after the MCP was released, but I don't worry too much - in no time it will take over Blender just like Cursor is taking over VSCode.

Rishi Ajith 的头像
Rishi Ajith1 年前

It looks like 💩 tbh. A person could have made it faster and better than whatever this is.

Edgy raven 的头像
Edgy raven1 年前

honestly, with so much texting around, you might as well learn the techniques. I use Blender, it's not exactly rocket science.

Emm 🔜 GDC 2025 的头像
Emm 🔜 GDC 20251 年前

The text is from the model (Claude) not me 🫣

XCecil 的头像
XCecil1 年前

I’ve been avoiding Claud because it feels like there’s too much advertising which means it must not be free.

Transfigured Human D (parody) 的头像
Transfigured Human D (parody)1 年前

Waste of time, it looks like ass and the model looks like ass

Myonisto 🟧 ᛤ 的头像
Myonisto 🟧 ᛤ1 年前

This looks awesome, do you know how i can use cursor instead of claude

Emm 🔜 GDC 2025 的头像
Emm 🔜 GDC 20251 年前

Check out @VisionaryxAI is doing it. It might work better tbh.

相关视频

3D-LLM: Injecting the 3D World into Large Language Models paper page: Large language models (LLMs) and Vision-Language Models (VLMs) have been proven to excel at multiple tasks, such as commonsense reasoning. Powerful as these models can be, they are not grounded in the 3D physical world, which involves richer concepts such as spatial relationships, affordances, physics, layout, and so on. In this work, we propose to inject the 3D world into large language models and introduce a whole new family of 3D-LLMs. Specifically, 3D-LLMs can take 3D point clouds and their features as input and perform a diverse set of 3D-related tasks, including captioning, dense captioning, 3D question answering, task decomposition, 3D grounding, 3D-assisted dialog, navigation, and so on. Using three types of prompting mechanisms that we design, we are able to collect over 300k 3D-language data covering these tasks. To efficiently train 3D-LLMs, we first utilize a 3D feature extractor that obtains 3D features from rendered multi- view images. Then, we use 2D VLMs as our backbones to train our 3D-LLMs. By introducing a 3D localization mechanism, 3D-LLMs can better capture 3D spatial information. Experiments on ScanQA show that our model outperforms state-of-the-art baselines by a large margin (e.g., the BLEU-1 score surpasses state-of-the-art score by 9%). Furthermore, experiments on our held-in datasets for 3D captioning, task composition, and 3D-assisted dialogue show that our model outperforms 2D VLMs. Qualitative examples also show that our model could perform more tasks beyond the scope of existing LLMs and VLMs.

AK

249,798 次观看 • 3 年前