Загрузка видео...
Не удалось загрузить видео
Video understanding isn't just recognizing —it demands reasoning across thousands of frames. Meet Long-RL🚀 Highlights: 🧠 Dataset: LongVideo-Reason — 52K QAs with reasoning. ⚡ System: MR-SP - 2.1× faster RL for long videos. 📈 Scalability: Hour-long videos (3,600 frames) RL on a single node (8×A100s). 🖼️📝🎵 RL training for... show more
31,855 просмотров • 1 год назад •via X (Twitter)
Комментарии: 3

andrea panizza1 год назад
Really nice of you to release the training code, thanks! Why not releasing also the final weights, though?

Yukang Chen1 год назад
Here is the weight, thanks.

andrea panizza1 год назад
Thank you for releasing them! Great job 🙏
