Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Introducing Omni, one unified model can support any-to-any multimodal modeling, including multimodal understanding, image/video generation and editing, world modeling and 3D reconstruction. All in one that adopts standard mixture-of-experts arch with only 3B activations.

32,674 görüntüleme • 5 ay önce •via X (Twitter)

9 Yorum

Ceyuan Yang profil fotoğrafı
Ceyuan Yang5 ay önce

Any-to-any training enables multimodal context unrolling where the model explicitly reasons across multiple modal representations before producing predictions. This reframes “unification” as a mechanism that scales context—not only in length, but in structure and utility for downstream decisions.

Ceyuan Yang profil fotoğrafı
Ceyuan Yang5 ay önce

Started last summer, but shared much later and less detailed than we hoped. It was built by a small team under many constraints, with plenty of imperfections, but also with a lot of support from people who truely believed in it. It only made us believe more: the future is multimodal.

Ceyuan Yang profil fotoğrafı
Ceyuan Yang5 ay önce

Homepage: ArXiv:

Kangfu Mei profil fotoğrafı
Kangfu Mei5 ay önce

Congratulations Ceyuan! That’s great work!

Jiazhi Yang profil fotoğrafı
Jiazhi Yang5 ay önce

Congrats Ceyuan!

P.MOHANSRINIVAS profil fotoğrafı
P.MOHANSRINIVAS5 ay önce

Any GitHub repo link @CeyuanY for testing results

Leo@Yuhao profil fotoğrafı
Leo@Yuhao5 ay önce

Congrats

Astrid Wilde 🌞 profil fotoğrafı
Astrid Wilde 🌞5 ay önce

interesting

Kairun Wen profil fotoğrafı
Kairun Wen5 ay önce

Congrats🥳!

Benzer Videolar