📢📢📢Introducing xGen-MM-Vid (BLIP-3-Video)! This highly efficient multimodal language model...

Salesforce AI Research's profile picture

Salesforce AI Research

12,550 görüntüleme • 1 yıl önce

📢Introducing Generated Reality📢 A world model for XR that...

Gordon Wetzstein's profile picture

Gordon Wetzstein

20,045 görüntüleme • 6 ay önce

Apparently SeeDance 2 is down for the time being...

TomLikesRobots🤖's profile picture

TomLikesRobots🤖

50,451 görüntüleme • 7 ay önce

NVIDIA just released a very impressive text-to-video paper. Video...

Lior Alexander's profile picture

Lior Alexander

158,600 görüntüleme • 3 yıl önce

The Hidden Language of Diffusion Models paper page: tackle...

AK's profile picture

AK

41,830 görüntüleme • 3 yıl önce

introducing FLUX 3. beyond video, this is the first...

Krea's profile picture

Krea

32,550 görüntüleme • 1 ay önce

📢 Good Monday BORG LOVE Morning, awesome BORGERS and...

SwissBorg's profile picture

SwissBorg

12,847 görüntüleme • 1 yıl önce

🚀New paper out - We present Video-MSG (Multimodal Sketch...

Jialu Li's profile picture

Jialu Li

35,060 görüntüleme • 1 yıl önce

Most recent diffusion language model research (that I’ve seen)...

nathan (in sf)'s profile picture

nathan (in sf)

40,440 görüntüleme • 7 ay önce

Grok Imagine, the AI video model that doesn't want...

Invideo's profile picture

Invideo

1,137,773 görüntüleme • 5 ay önce

A major question in multimodal modeling is how to...

Amir Zamir's profile picture

Amir Zamir

23,927 görüntüleme • 2 ay önce

LTX 2.3 is now live on Krea. this video...

Krea's profile picture

Krea

21,845 görüntüleme • 4 ay önce

Wonderland: Navigating 3D Scenes from a Single Image Contributions:...

MrNeRF's profile picture

MrNeRF

52,849 görüntüleme • 1 yıl önce

Fine-tune DeepSeek-OCR on your own language! (100% local) DeepSeek-OCR...

Akshay 🚀's profile picture

Akshay 🚀

126,213 görüntüleme • 10 ay önce

📢 Our xData Godin reply library just got an...

DIN⏳'s profile picture

DIN⏳

58,606 görüntüleme • 2 yıl önce

✨ Every time the video models get better, the...

@levelsio's profile picture

@levelsio

334,317 görüntüleme • 1 yıl önce

We got early access to the Gemini Omni API...

Hyperagent's profile picture

Hyperagent

325,111 görüntüleme • 2 ay önce

The recent Massachusetts Institute of Technology (MIT) CSAIL paper...

ChainGPT's profile picture

ChainGPT

82,385 görüntüleme • 7 ay önce