正在加载视频...

视频加载失败

Introducing 👀Stereo4D👀 A method for mining 4D from internet stereo videos. It enables large-scale, high-quality, dynamic, *metric* 3D reconstructions, with camera poses and long-term 3D motion trajectories. We used Stereo4D to make a dataset of over 100k real-world 4D scenes.

94,683 次观看 • 1 年前 •via X (Twitter)

16 条评论

Linyi Jin 的头像
Linyi Jin1 年前

This type of data is ideal for learning the structure and dynamics of the real world. We gave this a shot — extending DUSt3R to model 3D motion, and training on our dataset. Given a pair of frames, our model predicts a 3D point cloud, and corresponding 3D motion trajectories.

Linyi Jin 的头像
Linyi Jin1 年前

See more scenes & details of how it works on our website: Paper: Thanks to the great team! Richard Tucker, @zhengqi_li, David Fouhey, @Jimantha, @holynski_ Please stay tuned for updates on data & code.

Sherwin Bahmani 的头像
Sherwin Bahmani1 年前

Congrats, will be super useful for the 4D reconstruction/generation community!!

Chen Wang 的头像
Chen Wang1 年前

Very amazing work! Wondering how to predict the 3D point trajectory of a video after pairwise prediction of DynaDust3r?

Linyi Jin 的头像
Linyi Jin1 年前

Thanks Chen! We've tried using DynaDust3r to predict motion at any time between two frames. We haven't tried to extend it to take all frames of a video, but could be a cool future direction.

Pedro Milcent 的头像
Pedro Milcent1 年前

Excellent paper @jin_linyi ! Exciting potential for models leveraging multi-modal data 🦾

IceTTT 的头像
IceTTT1 年前

Very interesting work!

Dejvuzzz 的头像
Dejvuzzz1 年前

This should be Good for 3D tracking . @jin_linyi . both 3D camera tracking and 3D mesh tracking.

Feiyu Yao 的头像
Feiyu Yao1 年前

great work! when will it be released? can't wait to use it.

☀️ Leon-Gerard Vandenberg 🇺🇸 🇳🇱 🇨🇦 🇦🇺 的头像
☀️ Leon-Gerard Vandenberg 🇺🇸 🇳🇱 🇨🇦 🇦🇺1 年前

🤩 wow 🤩

☀️ Leon-Gerard Vandenberg 🇺🇸 🇳🇱 🇨🇦 🇦🇺 的头像
☀️ Leon-Gerard Vandenberg 🇺🇸 🇳🇱 🇨🇦 🇦🇺1 年前

The world simulation and the singularity is near @PeterDiamandis @salimismail

Nikodeam.eth 的头像
Nikodeam.eth1 年前

Me patiently waiting for code

Time is Light 的头像
Time is Light1 年前

@zhengqi_li Time is Light !

HU Wenbo 的头像
HU Wenbo1 年前

Cool!

Maya N 的头像
Maya N1 年前

I'm blown away by the scale of your dataset! How do you plan to make it available for others to use?

Bilawal Sidhu 的头像
Bilawal Sidhu1 年前

Really cool to see VR180 videos have this impact -- this'll be huge for the computer vision community!

相关视频