Video yükleniyor...
Video Yüklenemedi
Variable-length compressive tokenization is a promising direction. Not just for efficiency, but for actually learning powerful representations. FlexTok scratched the surface of this direction with images. But real-world data has additional structure that can be tapped, such as the temporal structure of videos. VideoFlexTok develops this concept for video,... show more
17,775 görüntüleme • 4 ay önce •via X (Twitter)
1 Yorum

Sivan Doveh4 ay önce
Great paper Amir! Super interesting
