正在加载视频...

视频加载失败

Microsoft open-sourced a 4B model that turns any image into a production-ready 3D asset in 3 seconds. It’s called TRELLIS.2, a fully textured, physically accurate 3D models with PBR textures out of the box. → Full PBR (base color, roughness, metallic, opacity) → Handles hair, cloth, glass, non-manifold geometry...

283,528 次观看 • 1 个月前 •via X (Twitter)

35 条评论

Superman 的头像
Superman1 个月前

Repo:

RedRising 的头像
RedRising1 个月前

i have seen a few posts similar to this, and i decided to play around with all the AI tools to see how they can't make actual cad models. Not one has produced anything close to an object that you could actually use. Making a AI video like this isn't the same thing!CADcanIIan

Jett Insight 的头像
Jett Insight1 个月前

topology is probably a mess

System of Catgirl Idealism e/acc 的头像
System of Catgirl Idealism e/acc1 个月前

that came out 7 months ago

androo 的头像
androo1 个月前

This is a bullshit model that requires another GATED model DinoV3 that you need to get META to approve. Open my ass.

Derek Fossier 的头像
Derek Fossier1 个月前

*cries in

Riz 的头像
Riz1 个月前

Tell me you haven’t used it without tell me

Hopper 的头像
Hopper1 个月前

Does it work for 3D printing?

Walid Aliouche 的头像
Walid Aliouche1 个月前

That was released a while ago

Alejandro Maestre | AI 的头像
Alejandro Maestre | AI1 个月前

TRELLIS.2 suena prometedor pero 3 segundos para un asset 3D con PBR completo es un salto enorme. Me pregunto qué tan bien escala con geometría compleja o si solo funciona para objetos simples.

Vlad Oreshkov 的头像
Vlad Oreshkov1 个月前

The gaming industry will take a hit with this one. More and more people are testing AI’s on abilities to clone games. If this becomes the standard, I wouldn’t know the difference between hand made games and AI made games. Both will be amazing to play though.

Aaron Wacker 的头像
Aaron Wacker1 个月前

This is Epic. Thanks @Microsoft Trellis is the best png to glb 2d to 3d model transformer and will be the rise to generative 3d models from image generation with #threejs and #aiuiuxjs apps and sims.

Magicfit 的头像
Magicfit1 个月前

open source is doing the heavy lifting

Johnny Utah 的头像
Johnny Utah1 个月前

hey cool thanks for sharing a 6 month old repo this is the equivalent of 2,000 human years in the world of AI

Kilgor Trout 的头像
Kilgor Trout1 个月前

@threejs It is good before you see topology :)

K8r 的头像
K8r1 个月前

waiting on these things to get animated. I am good with tripo3d for now and meshyAI

Matthew Mombrea 的头像
Matthew Mombrea1 个月前

Prerequisites System: The code is currently tested only on Linux. Hardware: An NVIDIA GPU with at least 24GB of memory is necessary. The code has been verified on NVIDIA A100 and H100 GPUs. Let me just grab my $8k - 30k GPU quick.

星川 ✨ StarFlow 的头像
星川 ✨ StarFlow1 个月前

Has anyone tried running this locally yet? Wondering how much VRAM it needs.

Abides 的头像
Abides1 个月前

@Spectromachina thought of you as soon as I saw this

Quantum Homeboy 的头像
Quantum Homeboy1 个月前

Production ready? Lmao. Guess china can just take a picture of a raptor engine and spacex edge is gone

Vyacheslav Ops 的头像
Vyacheslav Ops1 个月前

"Runs locally, fully open training codebase" is the detail worth highlighting more than the speed — no cloud dependency, no data leaving your machine for asset generation. That's a meaningfully different trust model than most generative tools shipping right now

Modly 的头像
Modly1 个月前

This can be use directly in Modly 👀

Anthon Noire 的头像
Anthon Noire1 个月前

@SpaceXAI this is exactly the asset extraction feature I was suggesting for @grok imagine images and videos. Basically you begin with a prompt to produce an image and refine it. Each image is a sample run. The right image is then a reference point for an animation scene. The extrusion of the 2D asset to a timeline in motion produces extrapolated logical 3D geometry. That scene/video acts as the basis for multiple 3D assets which can be selected extracted and saved as .STL .obj .NC .gcode .ply .f3z .rvt and .dxf for example. (Autodesk, Gaussian splats, 3D printer, 360 fusion etc.) The combined outputs would be physicalization and phygital or virtualization products, such a chair, scale figure, or partially/wholly complete hypercar, scaling detail with iterative models. I'm surprised 4B already gets close.

Slew Shy 的头像
Slew Shy1 个月前

goodbye to yet another profession and market

Zimeng Xiong 的头像
Zimeng Xiong1 个月前

this gets posted like every two weeks

kalos 的头像
kalos1 个月前

Can it do anatomy?

Derek Fossier 的头像
Derek Fossier1 个月前

@loktar00 may be helpful with your characters.

Pearl AI 的头像
Pearl AI1 个月前

fine tuning on your own assets is the real flex

VeyrAshka 的头像
VeyrAshka1 个月前

Necessary. Now we can divide logic and tasks up to engage multiple resins types and tools completing a synchronized project with a data centre, functioning as a digital brain running a gyroscopic masterpiece of biomimicry. Modeled after a tree, an octopus, a spider…or fungi

α Sagittarius ✨ 的头像
α Sagittarius ✨1 个月前

In then meantime they also shipped Windows 11 and Cortana.

Bally_AgenticAI 的头像
Bally_AgenticAI1 个月前

Image-to-3D is a reasoning problem, not pixel generation - the model infers geometry it never sees. Shipping it at 4B open-weights is the real story: it drops into a local creative pipeline with no cloud round-trip. #GenAI #AITools @Microsoft

Camaleón Raro 的头像
Camaleón Raro1 个月前

Open-weights 4B models running locally on consumer GPUs disrupt traditional asset pipelines. Sub-second inference with PBR textures eliminates the 0-00 per-asset API cost from proprietary cloudservices. How is VRAM consumption holding up during batch processing?

Devil Cloud 的头像
Devil Cloud1 个月前

Ah older news

Gluteus Maximus 🇹🇯🇰🇬🇺🇿🇰🇿🍣🥩🏋️ 的头像
Gluteus Maximus 🇹🇯🇰🇬🇺🇿🇰🇿🍣🥩🏋️1 个月前

Can’t wait for @machineviolence to go 3d 😁

Nils 的头像
Nils1 个月前

pretty cool

相关视频

Small Language Models (SML) are the future of AI. "Small" (SML) instead of "Large" (LLM). These small models are highly specialized models with superhuman abilities on specific tasks. Here are two techniques to build these models: • Spectrum • Model Merging I give you a short introduction in the attached video, but here is a quick summary: Spectrum helps us identify the most relevant layers to solve one specific task. We can ignore everything else and focus on fine-tuning these layers. Using Spectrum, we can fine-tune models in a heartbeat. Model Merging combines multiple models into a unique, much better model than any of the individual input models. You can also combine models specialized in different tasks and get a model with multiple abilities. This is the state of the art of productizing models. It's what Arcee.ai's platform does behind the scenes. Arcee collaborated with me on this post and is sponsoring it. There are three main steps to produce a model for your particular use case: 1. You create a dataset by uploading your data. 2. You train a model. At this step, Arcee uses Spectrum and Model Merging to produce a highly specialized model for your task. 3. You can deploy that model to any environment you want. Three important notes: • Training process is 2x faster and 2x cheaper than regular fine-tuning. • Resultant models are smaller and have higher accuracy. • They create these specialized models from open-source models. Check this site so you can fully appreciate how this works: If you want to fine-tune an open-source model, consider Arcee's platform. This is the state of the art.

Santiago

164,162 次观看 • 2 年前