Diffusion models have amazing image creation abilities. But how...

Alex Li's profile picture

Alex Li

95,304 views • 2 years ago

1/ Happy to share UniDisc - Unified Multimodal Discrete...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

104,934 views • 1 year ago

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,052 views • 1 year ago

We introduce TurboEdit -- simple text-based image editing in...

Richard Zhang's profile picture

Richard Zhang

39,306 views • 1 year ago

High-resolution image and video generation is hitting a wall...

Gordon Wetzstein's profile picture

Gordon Wetzstein

164,096 views • 4 months ago

Chop the gradients ✂️! We found that truncating decoder...

Felix Heide's profile picture

Felix Heide

28,323 views • 3 months ago

We are also releasing self-contained lecture notes that explain...

Peter Holderrieth's profile picture

Peter Holderrieth

475,659 views • 4 months ago

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 views • 2 years ago

Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering discuss: The...

AK's profile picture

AK

19,101 views • 1 year ago

Diffusion models are great, but we can squeeze out...

Yossi Gandelsman's profile picture

Yossi Gandelsman

34,863 views • 4 months ago

🚀 Self-speculation brings 6.75x real speedup for LLM generation...

Pavlo Molchanov's profile picture

Pavlo Molchanov

66,554 views • 2 months ago

NVIDIA just released a very impressive text-to-video paper. Video...

Lior Alexander's profile picture

Lior Alexander

158,565 views • 3 years ago

Can you make a jigsaw puzzle with two different...

Daniel Geng's profile picture

Daniel Geng

125,806 views • 2 years ago

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,123 views • 3 years ago

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 views • 1 year ago

Create a 3D model from a single image, set...

Bilawal Sidhu's profile picture

Bilawal Sidhu

92,792 views • 2 years ago

🚀 Introducing GenLit – Reformulating Single-Image Relighting as Video...

Haven Feng's profile picture

Haven Feng

22,442 views • 1 year ago

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,746 views • 3 years ago

🤔 How to fine-tune an Imitation Learning policy (e.g.,...

Tongzhou Mu 🤖🦾🦿's profile picture

Tongzhou Mu 🤖🦾🦿

16,959 views • 1 year ago

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 views • 1 year ago

🎬Comparing AI video models with no prompt: • Gen-3...

Heather Cooper's profile picture

Heather Cooper

29,138 views • 1 year ago