Video yükleniyor...
Video Yüklenemedi
How well do egocentric human videos scale? Apparently, really well. Dyna Robotics has introduced Dyna-2, a world-action model (WAM) pre-trained on more than 1 million hours of real egocentric human video. It learns by jointly predicting the next video frames and the actions, and that video prediction (world modeling)... show more
15,970 görüntüleme • 1 ay önce •via X (Twitter)
6 Yorum

Sneak peek at Dyna's more refined humanoid platform.

One million hours of human data: 43.8M clips 97,160 unique task instructions 9,917 distinct objects "Video co-training is the primary driver for establishing cross-embodiment transfer scaling law; it is both necessary and sufficient for the transfer to scale with data."

This is a huge signal for robotics 🤖 The fact that human video can scale pre-training across different robot embodiments with zero robot data is especially impressive. If this trend holds, the path from human demonstrations to capable robots could get dramatically faster. 🔥

Good to see a startup shipping real results. 87% vs 46% at unseen sites is the difference between a demo and a product.

nice

WAMs are super cool We need a benchmark called “we dropped the robot into a random 30-year-old factory on Tuesday, how long until it works”
