Introducing Diffusion Forcing, which unifies next-token prediction (eg LLMs)...

Boyuan Chen's profile picture

Boyuan Chen

208,868 views • 2 years ago

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 views • 1 year ago

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 views • 2 years ago

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 views • 1 year ago

Most recent diffusion language model research (that I’ve seen)...

nathan (in sf)'s profile picture

nathan (in sf)

40,440 views • 7 months ago

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 views • 1 year ago

You can't 3D reconstruct glass from images... ...WRONG! Thanks...

Jonathan Stephens's profile picture

Jonathan Stephens

17,712 views • 8 months ago

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,160 views • 3 years ago

Google dropped a new AI paper called LUMIERE. It's...

Bilawal Sidhu's profile picture

Bilawal Sidhu

44,822 views • 2 years ago

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 views • 1 year ago

Clarity Upscaler now works with Flux! 🥳 Clarity started...

philz1337x's profile picture

philz1337x

71,216 views • 1 year ago

The theory of higher order topological dynamics, which combines...

Maurizio Iβλἄ's profile picture

Maurizio Iβλἄ

40,749 views • 1 year ago

🎥 Video generation is hitting the memory wall. As...

Haocheng Xi's profile picture

Haocheng Xi

65,008 views • 4 months ago

Here are more results from #RigidFormer: predicting physical dynamics...

Zhiyang (Frank) Dou's profile picture

Zhiyang (Frank) Dou

21,837 views • 3 months ago

introducing the media synthesis museum an active and interactive...

apolinario (poli)'s profile picture

apolinario (poli)

11,143 views • 2 months ago

Added context to my tiny diffusion model to enable...

Nathan Barry's profile picture

Nathan Barry

89,040 views • 10 months ago

Meet Stable Audio 3.0, the open-weight model family built...

Stability AI's profile picture

Stability AI

166,625 views • 3 months ago

Today Terraport passes a huge milestone in its journey...

Terraport Finance's profile picture

Terraport Finance

30,497 views • 2 years ago

Depth Any Video with Scalable Synthetic Data AI physicists...

MrNeRF's profile picture

MrNeRF

27,428 views • 1 year ago