1/ Happy to share UniDisc - Unified Multimodal Discrete...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

104,934 Aufrufe • vor 1 Jahr

1/ Happy to share VADER: Video Diffusion Alignment via...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

13,368 Aufrufe • vor 1 Jahr

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 Aufrufe • vor 1 Jahr

Today, we are releasing Stable Video Diffusion, our first...

Stability AI's profile picture

Stability AI

1,024,532 Aufrufe • vor 2 Jahren

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,052 Aufrufe • vor 1 Jahr

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 Aufrufe • vor 1 Jahr

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 Aufrufe • vor 2 Jahren

Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering discuss: The...

AK's profile picture

AK

19,101 Aufrufe • vor 1 Jahr

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,123 Aufrufe • vor 3 Jahren

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 Aufrufe • vor 1 Jahr

Chop the gradients ✂️! We found that truncating decoder...

Felix Heide's profile picture

Felix Heide

28,399 Aufrufe • vor 3 Monaten

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 Aufrufe • vor 1 Jahr

We took a 30B model and split it in...

NVIDIA AI's profile picture

NVIDIA AI

759,419 Aufrufe • vor 1 Monat

🚀 Self-speculation brings 6.75x real speedup for LLM generation...

Pavlo Molchanov's profile picture

Pavlo Molchanov

66,554 Aufrufe • vor 2 Monaten

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,746 Aufrufe • vor 3 Jahren

🇨🇳 Another great Chinese Model, OmniHuman-1.5 from ByteDance Turns...

Rohan Paul's profile picture

Rohan Paul

63,859 Aufrufe • vor 11 Monaten

We release Diamond Maps💎 unlocking accurate and efficient guidance...

Peter Holderrieth's profile picture

Peter Holderrieth

60,179 Aufrufe • vor 3 Monaten

Video diffusion models have strong implicit representations of 3D...

Michael Black's profile picture

Michael Black

22,182 Aufrufe • vor 7 Monaten

You can't 3D reconstruct glass from images... ...WRONG! Thanks...

Jonathan Stephens's profile picture

Jonathan Stephens

17,712 Aufrufe • vor 7 Monaten

Break-A-Scene: Extracting Multiple Concepts from a Single Image introduce...

AK's profile picture

AK

154,511 Aufrufe • vor 3 Jahren

We’ve seen humanoid robots walk around for a while,...

Yanjie Ze's profile picture

Yanjie Ze

75,271 Aufrufe • vor 1 Jahr