🚨Announcing our #ICLR2025 Oral! 🔥Diffusion LMs are on the...

Marianne Arriola @ICML's profile picture

Marianne Arriola @ICML

150,079 views • 1 year ago

Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models...

Tanishq Mathew Abraham, Ph.D.'s profile picture

Tanishq Mathew Abraham, Ph.D.

21,813 views • 1 year ago

MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers paper...

AK's profile picture

AK

25,449 views • 2 years ago

Switti -- a new scale-wise transformer for text-to-image generation...

Gradio's profile picture

Gradio

29,314 views • 1 year ago

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 views • 2 years ago

1/ Happy to share UniDisc - Unified Multimodal Discrete...

Mihir Prabhudesai's profile picture

Mihir Prabhudesai

104,934 views • 1 year ago

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 views • 1 year ago

🚀 Self-speculation brings 6.75x real speedup for LLM generation...

Pavlo Molchanov's profile picture

Pavlo Molchanov

66,604 views • 2 months ago

Chop the gradients ✂️! We found that truncating decoder...

Felix Heide's profile picture

Felix Heide

28,399 views • 4 months ago

🎥 Video generation is hitting the memory wall. As...

Haocheng Xi's profile picture

Haocheng Xi

65,008 views • 3 months ago

Today, we are releasing Stable Video Diffusion, our first...

Stability AI's profile picture

Stability AI

1,024,598 views • 2 years ago

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,830 views • 3 years ago

We took a 30B model and split it in...

NVIDIA AI's profile picture

NVIDIA AI

761,340 views • 1 month ago

Decentralized Diffusion Models power stronger models trained on more...

David McAllister's profile picture

David McAllister

46,415 views • 1 year ago

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 views • 1 year ago

I am blown away 🤯. Check this out! CameraCtrl...

MrNeRF's profile picture

MrNeRF

12,633 views • 1 year ago

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 views • 1 year ago

Here are more results from #RigidFormer: predicting physical dynamics...

Zhiyang (Frank) Dou's profile picture

Zhiyang (Frank) Dou

20,955 views • 3 months ago

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 views • 1 year ago

[SIGGRAPH ASIA '25] Detail-Enhanced Gaussian Splatting for Large-Scale Volumetric...

MrNeRF's profile picture

MrNeRF

42,456 views • 9 months ago

Most recent diffusion language model research (that I’ve seen)...

nathan (in sf)'s profile picture

nathan (in sf)

40,440 views • 7 months ago