Загрузка видео...

Не удалось загрузить видео

На главную

🚀Introducing LLaVA-NeXT Interleave: Now AI can understand and reason with multiple images at once - This opens up multi-image scenarios like multi-frame videos, multi-view 3D, and multiple inter-leaved images. - An all round LMM that can understand videos, images, and 3D More⬇️

27,665 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 8

Фото профиля Gradio
Gradio2 лет назад

LLaVA-NeXT-Interleave🔥 - Interleave data format unifies different tasks. - New datasets on 🤗Hub: 1️⃣M4-Instruct, high-quality dataset, 1.1M samples from domains: multi-image, video, 3D & single-image 2️⃣LLaVA-Interleave Bench - Set of tasks to evaluate multi-image capabilities

Фото профиля Gradio
Gradio2 лет назад

LLaVA-NeXT-Interleave💪 - Attached videos show how it can explain jokes and understand content spread in multiple images and videos 🤯 - SoTA Performance, both, in multi and single images - Matches in perf with LLaVA-NeXT - Improved performance in video tasks

Фото профиля Gradio
Gradio2 лет назад

Gradio Multimodal Demo for LLaVA-NeXT-Interleave😍 : Models and Datasets are on 🤗 Hub:

Фото профиля Stark
Stark2 лет назад

how to finetune?

Фото профиля Omri Kaduri
Omri Kaduri2 лет назад

How can you refer to the order of the images in the prompt? Simply saying "first image" is enough? Like -"is the object in the first image shown in the second image"

Фото профиля Lily Zhang
Lily Zhang2 лет назад

How does it understand 3D?

Фото профиля Gradio
Gradio2 лет назад

Different views as multiple image input

Фото профиля Gradio
Gradio2 лет назад

Love this! 💡

Похожие видео