Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

🚀Exicited to present our new work "InvertAvatar: Incremental GAN Inversion for Generalized Head Avatars". Given single or more source images, our method rapidly reconstructs photorealistic 3D facial avatars under 1s. Paper: Project page comes soon!

16,741 Aufrufe • vor 2 Jahren •via X (Twitter)

10 Kommentare

Profilbild von Jingxiang Sun
Jingxiang Sunvor 2 Jahren

Leveraging a 3D GAN prior, we introduce an effective recurrent inversion pipeline that aggregates temporal data and inherently supports flexible image inputs.

Profilbild von 田中義弘 | taziku CEO / AI × Creative
田中義弘 | taziku CEO / AI × Creativevor 2 Jahren

This is a promising technology. In Japan, there is a VTuber culture, so it will be greatly utilized!

Profilbild von Martin Shkreli (e/acc)
Martin Shkreli (e/acc)vor 2 Jahren

Nice job

Profilbild von Jingxiang Sun
Jingxiang Sunvor 2 Jahren

Thanks😄

Profilbild von Tonter_IAS
Tonter_IASvor 2 Jahren

Wow! 👏👏👏

Profilbild von BLENDER SUSHI 🫶 X - 24/7 Blenderian
BLENDER SUSHI 🫶 X - 24/7 Blenderianvor 2 Jahren

Pin the shoulders and it would be better.

Profilbild von Jingxiang Sun
Jingxiang Sunvor 2 Jahren

yeah that‘s exactly what we want to do next🤝

Profilbild von 210 Digital Marketing
210 Digital Marketingvor 2 Jahren

Nice 👍

Profilbild von wwwwg
wwwwgvor 2 Jahren

@Memdotai mem it

Profilbild von Mem
Memvor 2 Jahren

@JingxiangSun42 Saved! Here's the compiled thread: 🪄 AI-generated summary: "We are excited to present our new work, InvertAvatar, which rapidly reconstructs photorealistic 3D facial avatars from single or multiple source images in under 1 second....

Ähnliche Videos

📢📢 𝐀𝐯𝐚𝐭𝟑𝐫 📢📢 Avat3r creates high-quality 3D head avatars from just a few input images in a single forward pass with a new dynamic 3DGS reconstruction model. Video: Project: Our core idea is to make Gaussian Reconstruction Models animatable. We find that a simple cross-attention to an expression code sequence is already sufficient to model complex facial expressions. We then incorporate position maps from DUSt3R and feature maps from Sapiens to facilitate the prediction task. While DUSt3R's position maps act as a pixel-aligned initialization for the Gaussians' positions, the Sapiens feature maps help the cross-view transformer to match corresponding image tokens in the 4 input images. One major challenge in creating a 3D head avatar from smartphone images comes from inconsistent facial expressions when the subject could not remain perfectly static during the capture. We eliminate this static requirement by simply showing our model input images with different facial expressions during training. This technique makes our model robust to inconsistent input images later on. Finally, we show that despite the model has been trained with 4 input images, one can even create a 3D head avatar when only a single image is available. To achieve this, we employ a pre-trained 3D GAN to lift the single image to 3D and then render the 4 input images for our model. This allows us to create 3D head avatars from single images and even highly out-of-distribution examples like AI generated faces, paintings or statues. Great work by Tobias Kirschstein from his internship at Meta with Javier Romero, Artem Sevastopolsky, and Shunsuke Saito

Matthias Niessner

74,863 Aufrufe • vor 1 Jahr

🚀Announcing NeRSemble 3D Head Avatar Benchmark v2 Version 2 of the NeRSemble 3D Head Avatar Benchmark systematically evaluates several aspects of 3D head avatar creation. Our goal is to drive progress toward more realistic, robust, and generalizable avatar methods. 🔬Benchmark Tasks The NeRSemble Benchmark v2 features three core challenges: - Dynamic Novel View Synthesis - Monocular FLAME-driven Avatar Creation (updated) - Single-view 3D Face Reconstruction (new) 👉Explore the online leaderboard and submission system: 🆕What's new? 1. New Task: Single-view 3D Face Reconstruction Given a single portrait image, reconstruct an accurate 3D mesh either showing the input expression or a fully neutral one. Unlike prior benchmarks, the NeRSemble benchmark emphasizes diverse and challenging facial expressions, better reflecting real scenarios. For technical details, see the Pixel3DMM paper. 2. Updated task: Monocular FLAME-driven Avatar Creation We have improved the FLAME tracking that is used for both avatar creation from the monocular videos and avatar driving on the hidden test sequences. The updated benchmark task has: - more stable torso tracking - more expressive lip closures during speech - Improved mouth tracking for challenging facial expressions We hope that these improvements to the benchmark help drive the field forward. 🏆 CVPR 2026 Workshop & Prizes The NeRSemble benchmark will be featured at the CVPR 2026 Workshop on Photo-realistic 3D Head Avatars. Participants in the new and updated tasks have the opportunity to win: - 🎁RTX 5080 GPUs (sponsored by NVIDIA) - 🎤15-minute oral presentation at the workshop ⏰ Submission Deadline - May 26, 2026 Reach out to the amazing Tobias Kirschstein and Simon Giebenhain for more details :)

Matthias Niessner

30,098 Aufrufe • vor 6 Monaten

Nvidia announces GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning paper page: Gaussian splatting has emerged as a powerful 3D representation that harnesses the advantages of both explicit (mesh) and implicit (NeRF) 3D representations. In this paper, we seek to leverage Gaussian splatting to generate realistic animatable avatars from textual descriptions, addressing the limitations (e.g., flexibility and efficiency) imposed by mesh or NeRF-based representations. However, a naive application of Gaussian splatting cannot generate high-quality animatable avatars and suffers from learning instability; it also cannot capture fine avatar geometries and often leads to degenerate body parts. To tackle these problems, we first propose a primitive-based 3D Gaussian representation where Gaussians are defined inside pose-driven primitives to facilitate animation. Second, to stabilize and amortize the learning of millions of Gaussians, we propose to use neural implicit fields to predict the Gaussian attributes (e.g., colors). Finally, to capture fine avatar geometries and extract detailed meshes, we propose a novel SDF-based implicit mesh learning approach for 3D Gaussians that regularizes the underlying geometries and extracts highly detailed textured meshes. Our proposed method, GAvatar, enables the large-scale generation of diverse animatable avatars using only text prompts. GAvatar significantly surpasses existing methods in terms of both appearance and geometry quality, and achieves extremely fast rendering (100 fps) at 1K resolution.

AK

141,058 Aufrufe • vor 2 Jahren