How can we use test-time compute for spatial understanding?...

Ricardo Martin-Brualla's profile picture

Ricardo Martin-Brualla

21,685 次观看 • 1 年前

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 次观看 • 1 年前

Depth Any Video with Scalable Synthetic Data AI physicists...

MrNeRF's profile picture

MrNeRF

27,428 次观看 • 1 年前

New research from BlackForestLabsAI - Unofficial 🥳 Meet Self-Flow:...

Hila Chefer's profile picture

Hila Chefer

68,948 次观看 • 6 个月前

🔥Spatial intelligence requires world generation, and now we have...

Hong-Xing (Koven) Yu's profile picture

Hong-Xing (Koven) Yu

66,876 次观看 • 1 年前

(1/n) 🚀 With FastVideo, you can now generate a...

Hao AI Lab's profile picture

Hao AI Lab

78,660 次观看 • 1 年前

AI’s next frontier is Spatial Intelligence, a technology that...

Fei-Fei Li's profile picture

Fei-Fei Li

928,391 次观看 • 10 个月前

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos...

MrNeRF's profile picture

MrNeRF

24,729 次观看 • 1 年前

🤔Applying Depth Estimation models directly to videos can result...

Gradio's profile picture

Gradio

19,992 次观看 • 2 年前

(1/3) Can we turn text-to-image models into photorealistic 3D...

Matthias Niessner's profile picture

Matthias Niessner

34,779 次观看 • 2 年前

Generative models can’t discover what they can’t reach. We’re...

Riccardo De Santi's profile picture

Riccardo De Santi

40,016 次观看 • 10 天前

Sometimes I generate a video like this, and I...

Alex Patrascu's profile picture

Alex Patrascu

51,097 次观看 • 10 个月前

Introducing Kaleido💮 from AI at Meta — a universal...

Shikun Liu's profile picture

Shikun Liu

22,442 次观看 • 11 个月前

Can you make a jigsaw puzzle with two different...

Daniel Geng's profile picture

Daniel Geng

125,917 次观看 • 2 年前

1/ Happy to share VADER: Video Diffusion Alignment via...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

13,403 次观看 • 2 年前

Today, we are adding Stable Video Diffusion, our foundation...

Stability AI's profile picture

Stability AI

175,837 次观看 • 2 年前

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 次观看 • 1 年前

A major question in multimodal modeling is how to...

Amir Zamir's profile picture

Amir Zamir

23,927 次观看 • 1 个月前

Video understanding is the next frontier, but not all...

Saining Xie's profile picture

Saining Xie

210,896 次观看 • 1 年前

MRC is already deployed across all of OpenAI’s largest...

OpenAI's profile picture

OpenAI

146,115 次观看 • 4 个月前

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 次观看 • 2 年前