Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Introducing Omni, one unified model can support any-to-any multimodal modeling, including multimodal understanding, image/video generation and editing, world modeling and 3D reconstruction. All in one that adopts standard mixture-of-experts arch with only 3B activations.

32,674 Aufrufe • vor 5 Monaten •via X (Twitter)

9 Kommentare

Profilbild von Ceyuan Yang
Ceyuan Yangvor 5 Monaten

Any-to-any training enables multimodal context unrolling where the model explicitly reasons across multiple modal representations before producing predictions. This reframes “unification” as a mechanism that scales context—not only in length, but in structure and utility for downstream decisions.

Profilbild von Ceyuan Yang
Ceyuan Yangvor 5 Monaten

Started last summer, but shared much later and less detailed than we hoped. It was built by a small team under many constraints, with plenty of imperfections, but also with a lot of support from people who truely believed in it. It only made us believe more: the future is multimodal.

Profilbild von Ceyuan Yang
Ceyuan Yangvor 5 Monaten

Homepage: ArXiv:

Profilbild von Kangfu Mei
Kangfu Meivor 5 Monaten

Congratulations Ceyuan! That’s great work!

Profilbild von Jiazhi Yang
Jiazhi Yangvor 5 Monaten

Congrats Ceyuan!

Profilbild von P.MOHANSRINIVAS
P.MOHANSRINIVASvor 5 Monaten

Any GitHub repo link @CeyuanY for testing results

Profilbild von Leo@Yuhao
Leo@Yuhaovor 5 Monaten

Congrats

Profilbild von Astrid Wilde 🌞
Astrid Wilde 🌞vor 5 Monaten

interesting

Profilbild von Kairun Wen
Kairun Wenvor 5 Monaten

Congrats🥳!

Ähnliche Videos