Video yükleniyor...
Video Yüklenemedi
We found that to match the accuracy that S1 can achieve with just one example of prompting, current VLA models would need to be post-trained with 50-100 hours of data collection followed by fine-tuning!
295,652 görüntüleme • 8 gün önce •via X (Twitter)
0 Yorum
Yorum bulunmuyor
Orijinal gönderinin yorumları burada görünecek
