Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

AI just took 3D modeling to a whole new level 🤯 Introducing Neuralangelo, a new AI model by NVIDIA that reconstructs mind-blowingly detailed 3D surfaces directly from 2D videos — like photogrammetry on steroids. 🧙🏻‍♂️ Keep reading to see this crazy magic for yourself 🧵

721,064 Aufrufe • vor 3 Jahren •via X (Twitter)

11 Kommentare

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

So, what the heck is is this "photogrammetry" thing NVIDIA is supercharging with AI? TL;DR photogrammetry is the art & science of measuring stuff in the real world using images and other sensors (e.g. LiDAR). Here's a 60 second primer:

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

NVIDIA's new AI model is basically like photogrammetry on steroids. Why? Traditional photogrammetry can't handle repetitive structures, textureless surfaces or strong color variations. But Neuralangelo blends the tech behind Instant NeRF to capture every detail imaginable.

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Thus far, NeRFs & Photogrammetry have served different purposes: - NeRFs: Stunning visualizations (think flythroughs) but lack surface detail when turned into 3D meshes - Photogrammetry: Great for surface reconstruction (think measuring stuff) but not always visually appealing

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Neurangelo is the game-changer we've been waiting for. Its neural approach delivers the best of both worlds by bridging the gap between visuals and surface reconstruction. Say goodbye to blobby results: Neurangelo gives crisp 3D surfaces that make you salivate just seeing them.

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Neurangelo's coarse-to-fine optimization is like sculpting in 3D. First, a rough 3D scene emerges, just as a sculptor chisels a block and refines it bit by bit. Notice how details like the tree and bike rack may be missing at first, but they're successfully recovered later on.

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

In summary, Neuralangelo represents a significant advance: achieving both realistic visuals + finely detailed 3D models that stay true to surfaces, unlocking new possibilities for digital twins, gaming and visual effects. NVIDIA is back at it again!

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Plus, @nvidia's got 30 more AI projects debuting at the top computer vision conference (CVPR) later this month. So brace yourself! Follow me @bilawalsidhu for the latest in AI, AR and 3D tech. And if you found this helpful, feel free to like/retweet to show your support ❤️

Profilbild von J Dakota Powell
J Dakota Powellvor 3 Jahren

But someone has to retopologize all that geo for it to be even remotely usable - even if there's AI to do the retopo, inevitably someone has to step in to correct that version. Need reality checks on AI this-that, it's not a silver bullet.

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Agreed, nothing is ever a silver bullet, but much cleaner geometry is a big step in the right direction. I’ve frequently tweet that retopo needs some AI magic — seems asset optimization is the one thing humans are still better at (for now).

Profilbild von Rushik
Rushikvor 3 Jahren

Apple's XR tech coupled with this tech can "create new worlds"

Profilbild von Bilawal Sidhu
Bilawal Sidhuvor 3 Jahren

Indeed. And Apple does have their own photogrammetry API (which folks like Polycam and others use in their 3d capture apps) -- will be interesting to see if they announce any updates to it.

Ähnliche Videos

🔴 Finally! NVIDIA has finally made the code for Neuralangelo public! It has the ability to transform any video into a highly detailed 3D environment, and it's a technology related to but DIFFERENT from NeRF. 💡 Here's how it works: It takes a 2D video as input, showing an object, monument, building, landscape, etc., from various perspectives and analyzes details such as depth, size, and the shapes of objects. From this, the AI sketches an initial 3D model, similar to how an artist molds a figure. This representation is then refined to highlight more details, just as an artist would make the final touches when sculpting. The result is a 3D environment/model, perfect for use in any environment. Imagine the applications it will have for video games, cinema, virtual environments, VR, and more! 📽️🎮 💡 More details: A year ago, an article was presented on a groundbreaking technique called NVIDIA's Instant NeRF. This technique turns images into stunning 3D scenes in a short time, ideal for creating realistic models for video games and other applications. Although Instant NeRF had a lot of potential, the generated models were not perfect and often lacked detailed structures, appearing somewhat cartoonish. A year on, NVIDIA releases a new technique based on Instant NeRF, named Neuralangelo. This enhances the fidelity of surface structures. While NeRF reconstructs real objects in virtual environments from images or videos, Instant NeRF speeds up this process, and Neuralangelo further improves the quality, making the generated objects appear even more realistic when examined up close. Neuralangelo improves Instant NeRF's approach in two key ways related to the hash grid encoding technique: 1⃣ Numerical gradients have been used to compute higher-order derivatives as a smoothing operation. This optimizes the "hash grid" encoding using numerical rather than analytical gradients, providing a smoother input to the network that produces the 3D model. 2⃣ A "coarse-to-fine" optimization has been implemented in the hash grids to control different levels of detail. That is, they first focus on a smoothed version of the scene, and then refine it with more detailed updates. Well, as Arthur C. Clarke said, "Any sufficiently advanced technology is indistinguishable from magic."

Javi Lopez ⛩️

689,325 Aufrufe • vor 3 Jahren