Loading video...

Video Failed to Load

Go Home

Today, we're introducing SimFoundry, our real2sim2real framework at NVIDIA GEAR that automatically turns real-world scenes into simulation-ready worlds from a single image or video. Website: Paper: This work marks a major step for our team toward leveraging simulations and synthetic data for foundation model training and systematic policy evaluation...

80,828 views • 3 months ago •via X (Twitter)

15 Comments

Aetheris_Consulting 🇺🇸's profile picture
Aetheris_Consulting 🇺🇸3 months ago

Similar to this

Aetheris_Consulting 🇺🇸's profile picture
Aetheris_Consulting 🇺🇸3 months ago

Yuke awesome work here!!! Can you consider making a plug in for @UnrealEngine @unity @godotengine ? That was through a harness LLM can do physics computations paired with real2sim2real to do a number of things. This is such great work pls keep it up!!! 👍

Arthur Petron's profile picture
Arthur Petron3 months ago

How well do dynamics work? Like, if I drop the tennis ball from 0.5m in real does it bounce the same in sim?

Weijie Wang's profile picture
Weijie Wang2 months ago

Another antidote to hype—real frontier work with paper and open-source pledge, verifiable author; judge it by the arXiv paper, not the summary.

Sitarama Chekuri's profile picture
Sitarama Chekuri3 months ago

Looks amazing. Are the objects physics ready and interactive?

Mikhail's profile picture
Mikhail3 months ago

thanks habibi

Ferbin's profile picture
Ferbin3 months ago

Robot training always gets stuck building training environments. Automating that from photos changes the whole game.

Efstratios Gavves's profile picture
Efstratios Gavves3 months ago

@yukez Great work! I think it would be fair and nice to cite our DreMa work that was the first to introduce in ICLR 2025 the paradigm:

just10101's profile picture
just101012 months ago

Shameless @danfei_xu Stop stealing from students/ interviewees Shame on @gtcomputing

EB1A Experts's profile picture
EB1A Experts3 months ago

Impressive work. Bridging real-world scenes and simulation-ready environments from minimal input is an exciting step toward more scalable robotics and foundation model development. Looking forward to seeing the open-source release.

Larry Panozzo's profile picture
Larry Panozzo3 months ago

Wow, from one video from one lens with depth estimation…amazing. Most applications would use stereo I imagine? Results could be even better then, with a different version of this framework?

Zavian Pokharkar's profile picture
Zavian Pokharkar2 months ago

Yuke

Adel Dennaoui's profile picture
Adel Dennaoui2 months ago

When will the code be released? :)

AI Mastery Guide's profile picture
AI Mastery Guide3 months ago

Real2sim2real from a single image is huge for robotics training data, excited to see this open-sourced.

Kavya's profile picture
Kavya1 month ago

@yukez, when will you be releasing the code? Would love to try this out!

Related Videos

Most AI world models can generate beautiful scenes. Keeping those scenes alive for an hour without falling apart is the real challenge. That's what caught my attention about LingBot-World 2.0 (LingBot-World-Infinity) from Robbyant Instead of chasing longer videos, it focuses on something much harder: persistent, interactive worlds that stay coherent while you explore. A few highlights: • Generates worlds from a single frame and continuously responds to live user actions through a causal world model. • Streams stable 720p at 60 FPS in real time. The team reports a continuous 60 minute stress test across 20 different scenarios with no noticeable visual degradation. • Uses a Brain-Cerebellum co-simulation framework where a VLM plans events while the video model turns them into consistent world evolution. • Pilot and Director Agents help drive character behavior and introduce new objects and events. • Open sourced with a 14B flagship model, while the paper also describes a lightweight 1.3B version for a single consumer GPU. There is also an online interactive demo. The biggest takeaway? We're moving beyond AI that generates clips. We're getting closer to AI that generates living, evolving worlds you can actually interact with. And that feels like a much bigger shift than another jump in video quality. Explore more: 💻 Github: 🤗 Weights: 🌐 Website-with videos you can use : 🎮 Try it online: #Robbyant #LingBot #WorldModel #EmbodiedAI #OpenSource #Robotics #ad

Alif Khan

84,553 views • 2 months ago