Video wird geladen...
Video konnte nicht geladen werden
🤖 Introducing ω-0 (OMEGA-0) — a whole-body World Action Model for humanoids. Can one model make a humanoid walk, manipulate, and coordinate its whole body simultaneously across many real-world tasks? ω-0 does exactly that. ⚡ One unified model for multi-task whole-body loco-manipulation 🧠 Predicts future visual latents while generating... show more
18,940 Aufrufe • vor 1 Monat •via X (Twitter)
4 Kommentare

81.8% is the headline. Table 2 also lists ψ-0 at 44.5% on the same 11 tasks. That is a 37.3 point cap on this board. Progress sits at 90.3% while full success is 81.8%. The model often moves the task and still fails the success criterion. Those are two different scores. *table 2*

Terrified and wanting to put it on my Christmas wish list at the same time! 🤖😱🥳

in omega-0 does the predicted visual latent actually feed the action head at inference, or is it only there in the loss during training

81.8% across 11 household tasks is solid but the real question is what happens in environments that weren't in the training set. a warehouse with wet cardboard and shifting pallets is a different world than a home kitchen. does the world model generalize or does it memorize?
