Introducing Diffusion Forcing, which unifies next-token prediction (eg LLMs)...

Boyuan Chen's profile picture

Boyuan Chen

208,868 görüntüleme • 2 yıl önce

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 görüntüleme • 1 yıl önce

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 görüntüleme • 2 yıl önce

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 görüntüleme • 1 yıl önce

Most recent diffusion language model research (that I’ve seen)...

nathan (in sf)'s profile picture

nathan (in sf)

40,440 görüntüleme • 7 ay önce

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 görüntüleme • 1 yıl önce

You can't 3D reconstruct glass from images... ...WRONG! Thanks...

Jonathan Stephens's profile picture

Jonathan Stephens

17,712 görüntüleme • 8 ay önce

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,160 görüntüleme • 3 yıl önce

Google dropped a new AI paper called LUMIERE. It's...

Bilawal Sidhu's profile picture

Bilawal Sidhu

44,822 görüntüleme • 2 yıl önce

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 görüntüleme • 1 yıl önce

Clarity Upscaler now works with Flux! 🥳 Clarity started...

philz1337x's profile picture

philz1337x

71,216 görüntüleme • 1 yıl önce

The theory of higher order topological dynamics, which combines...

Maurizio Iβλἄ's profile picture

Maurizio Iβλἄ

40,749 görüntüleme • 1 yıl önce

🎥 Video generation is hitting the memory wall. As...

Haocheng Xi's profile picture

Haocheng Xi

65,008 görüntüleme • 4 ay önce

Here are more results from #RigidFormer: predicting physical dynamics...

Zhiyang (Frank) Dou's profile picture

Zhiyang (Frank) Dou

21,837 görüntüleme • 3 ay önce

introducing the media synthesis museum an active and interactive...

apolinario (poli)'s profile picture

apolinario (poli)

11,143 görüntüleme • 2 ay önce

Added context to my tiny diffusion model to enable...

Nathan Barry's profile picture

Nathan Barry

89,040 görüntüleme • 10 ay önce

Meet Stable Audio 3.0, the open-weight model family built...

Stability AI's profile picture

Stability AI

166,625 görüntüleme • 3 ay önce

Today Terraport passes a huge milestone in its journey...

Terraport Finance's profile picture

Terraport Finance

30,497 görüntüleme • 2 yıl önce

Depth Any Video with Scalable Synthetic Data AI physicists...

MrNeRF's profile picture

MrNeRF

27,428 görüntüleme • 1 yıl önce