How can we use test-time compute for spatial understanding?...

Ricardo Martin-Brualla's profile picture

Ricardo Martin-Brualla

21,685 Aufrufe • vor 1 Jahr

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 Aufrufe • vor 1 Jahr

Depth Any Video with Scalable Synthetic Data AI physicists...

MrNeRF's profile picture

MrNeRF

27,428 Aufrufe • vor 1 Jahr

New research from BlackForestLabsAI - Unofficial 🥳 Meet Self-Flow:...

Hila Chefer's profile picture

Hila Chefer

68,948 Aufrufe • vor 6 Monaten

🔥Spatial intelligence requires world generation, and now we have...

Hong-Xing (Koven) Yu's profile picture

Hong-Xing (Koven) Yu

66,876 Aufrufe • vor 1 Jahr

(1/n) 🚀 With FastVideo, you can now generate a...

Hao AI Lab's profile picture

Hao AI Lab

78,660 Aufrufe • vor 1 Jahr

AI’s next frontier is Spatial Intelligence, a technology that...

Fei-Fei Li's profile picture

Fei-Fei Li

928,391 Aufrufe • vor 10 Monaten

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos...

MrNeRF's profile picture

MrNeRF

24,729 Aufrufe • vor 1 Jahr

🤔Applying Depth Estimation models directly to videos can result...

Gradio's profile picture

Gradio

19,992 Aufrufe • vor 2 Jahren

(1/3) Can we turn text-to-image models into photorealistic 3D...

Matthias Niessner's profile picture

Matthias Niessner

34,779 Aufrufe • vor 2 Jahren

Generative models can’t discover what they can’t reach. We’re...

Riccardo De Santi's profile picture

Riccardo De Santi

40,016 Aufrufe • vor 11 Tagen

Sometimes I generate a video like this, and I...

Alex Patrascu's profile picture

Alex Patrascu

51,097 Aufrufe • vor 10 Monaten

Introducing Kaleido💮 from AI at Meta — a universal...

Shikun Liu's profile picture

Shikun Liu

22,442 Aufrufe • vor 11 Monaten

Can you make a jigsaw puzzle with two different...

Daniel Geng's profile picture

Daniel Geng

125,917 Aufrufe • vor 2 Jahren

1/ Happy to share VADER: Video Diffusion Alignment via...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

13,403 Aufrufe • vor 2 Jahren

Today, we are adding Stable Video Diffusion, our foundation...

Stability AI's profile picture

Stability AI

175,837 Aufrufe • vor 2 Jahren

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 Aufrufe • vor 1 Jahr

A major question in multimodal modeling is how to...

Amir Zamir's profile picture

Amir Zamir

23,927 Aufrufe • vor 1 Monat

Video understanding is the next frontier, but not all...

Saining Xie's profile picture

Saining Xie

210,896 Aufrufe • vor 1 Jahr

MRC is already deployed across all of OpenAI’s largest...

OpenAI's profile picture

OpenAI

146,115 Aufrufe • vor 4 Monaten

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 Aufrufe • vor 2 Jahren