正在加载视频...

视频加载失败

Super clean and efficient meshes by an AI? YES! The typical 3D Generative AI solutions produce lots of artifacts and usually way to many polygons due to volumetric approaches. In comparison “MeshGPT creates triangle meshes by autoregressively sampling from a transformer model that has been trained to produce tokens...

20,772 次观看 • 2 年前 •via X (Twitter)

9 条评论

René Schulte 的头像
René Schulte2 年前

No code (yet) but there's a project page for MeshGPT:

Hermes ᯅ 的头像
Hermes ᯅ2 年前

@RoszykAdam Very cool approach to the mesh generation challenge. These models look much closer to hand modeled 3D. I think this overcomes one of the biggest challenges, the ability to edit. With the dense meshes it’s very difficult to edit vs these simpler meshes.

René Schulte 的头像
René Schulte2 年前

@RoszykAdam You got it.

lewis ward 🇺🇦 🇺🇸 的头像
lewis ward 🇺🇦 🇺🇸2 年前

Nice!

teej 的头像
teej2 年前

@Doctor_Peepee we're so spot on

The Qubits Guy 的头像
The Qubits Guy2 年前

Now this is what I would like to see for Qubits. @Qubits_Toy

Primeclass AI 的头像
Primeclass AI2 年前

Woow! Super impressive!

Daniel Tabár 的头像
Daniel Tabár2 年前

misguided premise here... paper thin and pain in the ass polygons are terrible representations of reality. small ordered samples (voxels) all the way through is the way. see @atomontage_com

L Ξ Ο N V Λ N Κ Λ Μ Μ Ξ N @lvk@mastodon.online 的头像
L Ξ Ο N V Λ N Κ Λ Μ Μ Ξ N @[email protected]2 年前

very cool, I wonder how easy it is to link these scenes together using Like, in theory you could add an additional plane (door) with an src+href to another scene. Text to 3D worlds?

相关视频

📢📢 𝐀𝐯𝐚𝐭𝟑𝐫 📢📢 Avat3r creates high-quality 3D head avatars from just a few input images in a single forward pass with a new dynamic 3DGS reconstruction model. Video: Project: Our core idea is to make Gaussian Reconstruction Models animatable. We find that a simple cross-attention to an expression code sequence is already sufficient to model complex facial expressions. We then incorporate position maps from DUSt3R and feature maps from Sapiens to facilitate the prediction task. While DUSt3R's position maps act as a pixel-aligned initialization for the Gaussians' positions, the Sapiens feature maps help the cross-view transformer to match corresponding image tokens in the 4 input images. One major challenge in creating a 3D head avatar from smartphone images comes from inconsistent facial expressions when the subject could not remain perfectly static during the capture. We eliminate this static requirement by simply showing our model input images with different facial expressions during training. This technique makes our model robust to inconsistent input images later on. Finally, we show that despite the model has been trained with 4 input images, one can even create a 3D head avatar when only a single image is available. To achieve this, we employ a pre-trained 3D GAN to lift the single image to 3D and then render the 4 input images for our model. This allows us to create 3D head avatars from single images and even highly out-of-distribution examples like AI generated faces, paintings or statues. Great work by Tobias Kirschstein from his internship at Meta with Javier Romero, Artem Sevastopolsky, and Shunsuke Saito

Matthias Niessner

74,763 次观看 • 1 年前