Video wird geladen...
Video konnte nicht geladen werden
VideoLLaMA3, latest MLLMs for image and video understanding. 🖐️ 7B models: DocVQA: 94.9, MathVision: 26.2, VideoMME: 66.2/70.3, MLVU: 73.0 🤏 2B models for edge devices: MMMU: 45.3, VideoMME: 59.6/63.4 👊 Frontier-class video model with ONLY 3M video-text pairs
37,154 Aufrufe • vor 1 Jahr •via X (Twitter)
7 Kommentare

Gradiovor 1 Jahr
VideoLLaMA3 for Video understanding:

Gradiovor 1 Jahr
VideoLLaMA3 for Image understanding:

Gradiovor 1 Jahr
Learn more about VideoLLaMA3:

Rainmakervor 2 Jahren
Join me as I put several Machine Learning models head-to-head to see which one can beat the market and deliver strong returns. In this free Substack post I share several models that deliver better returns with much lower drawdown compared to Buy-and-Hold approach.

Discord Guysvor 1 Jahr
Wow, VideoLLaMA3 sounds like a beast! Can't wait to try it out.

Sujith Reddyvor 1 Jahr
@_akhaliq @Gradio are these models production level models? Scalable?

ヌルvor 1 Jahr
3M pairs is a small dataset for training.
