Introducing Diffusion Forcing, which unifies next-token prediction (eg LLMs)...

Boyuan Chen's profile picture

Boyuan Chen

208,868 Aufrufe • vor 2 Jahren

DimensionX: Create Any 3D and 4D Scenes from a...

MrNeRF's profile picture

MrNeRF

17,062 Aufrufe • vor 1 Jahr

Show-o One Single Transformer to Unify Multimodal Understanding and...

AK's profile picture

AK

124,048 Aufrufe • vor 2 Jahren

World modeling and imitation learning have largely been considered...

Abhishek Gupta's profile picture

Abhishek Gupta

11,430 Aufrufe • vor 1 Jahr

Most recent diffusion language model research (that I’ve seen)...

nathan (in sf)'s profile picture

nathan (in sf)

40,440 Aufrufe • vor 7 Monaten

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 Aufrufe • vor 1 Jahr

You can't 3D reconstruct glass from images... ...WRONG! Thanks...

Jonathan Stephens's profile picture

Jonathan Stephens

17,712 Aufrufe • vor 8 Monaten

Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation paper page:...

AK's profile picture

AK

375,160 Aufrufe • vor 3 Jahren

Google dropped a new AI paper called LUMIERE. It's...

Bilawal Sidhu's profile picture

Bilawal Sidhu

44,822 Aufrufe • vor 2 Jahren

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 Aufrufe • vor 1 Jahr

Clarity Upscaler now works with Flux! 🥳 Clarity started...

philz1337x's profile picture

philz1337x

71,216 Aufrufe • vor 1 Jahr

The theory of higher order topological dynamics, which combines...

Maurizio Iβλἄ's profile picture

Maurizio Iβλἄ

40,749 Aufrufe • vor 1 Jahr

🎥 Video generation is hitting the memory wall. As...

Haocheng Xi's profile picture

Haocheng Xi

65,008 Aufrufe • vor 4 Monaten

Here are more results from #RigidFormer: predicting physical dynamics...

Zhiyang (Frank) Dou's profile picture

Zhiyang (Frank) Dou

21,837 Aufrufe • vor 3 Monaten

introducing the media synthesis museum an active and interactive...

apolinario (poli)'s profile picture

apolinario (poli)

11,143 Aufrufe • vor 2 Monaten

Added context to my tiny diffusion model to enable...

Nathan Barry's profile picture

Nathan Barry

89,040 Aufrufe • vor 10 Monaten

Meet Stable Audio 3.0, the open-weight model family built...

Stability AI's profile picture

Stability AI

166,625 Aufrufe • vor 3 Monaten

Today Terraport passes a huge milestone in its journey...

Terraport Finance's profile picture

Terraport Finance

30,497 Aufrufe • vor 2 Jahren

Depth Any Video with Scalable Synthetic Data AI physicists...

MrNeRF's profile picture

MrNeRF

27,428 Aufrufe • vor 1 Jahr