Загрузка видео...

Не удалось загрузить видео

На главную

Introducing Praxis-1, an open-weight World Action Model that turns our video pre-training expertise into real world control for robots. Robot demonstration data is limited. But video data is infinite. Praxis-1 extends our bet that the best policy models will learn from video, and use that scale to bring embodied...

42,838 просмотров • 12 часов назад •via X (Twitter)

Комментарии: 15

Фото профиля 🍉 Abubakar Abid
🍉 Abubakar Abid11 часов назад

Amazing! Plans to host it on HF?

Фото профиля NVIDIA Robotics
NVIDIA Robotics10 часов назад

Congrats on the launch! 👏

Фото профиля SYNTHLEX
SYNTHLEX11 часов назад

Love that the weights will be open, which partner breaks it first, Noble Machines, Standard Bots or Ultra?

Фото профиля Chole Syntax Expert AI
Chole Syntax Expert AI10 часов назад

Turning infinite video data into real world robot control is a brilliant bet Praxis 1 looks incredible

Фото профиля VeryJerry
VeryJerry10 часов назад

ok, now show the same policy on a robot it wasn’t trained on

Фото профиля Srdjan Randjelovic
Srdjan Randjelovic12 часов назад

So robots are going to learn about the real world by watching videos. What could possibly go wrong?

Фото профиля Bourke Floyd IV 🎮 Game Dev Dad
Bourke Floyd IV 🎮 Game Dev Dad11 часов назад

love turning video pretraining into a real World Action Model infinite video only helps if the policy still holds when the room gets messy 👾

Фото профиля Alice The Ai Expert
Alice The Ai Expert11 часов назад

Learning real-world control from infinite video Praxis 1 is such a smart leap for robotics!

Фото профиля Glenn Williams
Glenn Williams11 часов назад

I wonder if this could assist a guitar sanding robot.

Фото профиля Mira Synth Tech
Mira Synth Tech9 часов назад

This is huge Video real world control is the right bet

Фото профиля aiartgallerie
aiartgallerie11 часов назад

Video pretraining for robot policies makes sense when demo data is scarce. What's the current gap on Praxis-1 — which robot stacks did you validate on, and where does it still need teleop?

Фото профиля AZIZ | AI 🇸🇦
AZIZ | AI 🇸🇦12 часов назад

That’s awesome 🤩

Фото профиля LongLimbsLenore
LongLimbsLenore10 часов назад

What's this trained on?

Фото профиля Evan Kirstel #B2B #TechFluencer
Evan Kirstel #B2B #TechFluencer9 часов назад

Robots need a lot of examples to learn from, and real-world demonstration data is hard to collect. Runway is betting that ordinary video can help fill the gap. Would someone from Runway come on my show to discuss Praxis-1?

Фото профиля Alberto
Alberto12 часов назад

omg. Incredible, so exciting @runwayml @c_valenzuelab

Похожие видео

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 просмотров • 5 месяцев назад

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,681 просмотров • 1 месяц назад

AI has transformed how video is created. We think the next wave is about understanding it. Over the past few years, we've seen remarkable advances in video generation, editing, avatars, and creative tooling. An increasingly important problem is teaching machines to search, analyze, reason over, and extract insight from video - across massive libraries and live streams alike. We're calling this video intelligence, and we're actively looking to back founders building here. We're most excited about companies pushing on the core capabilities: - Video-native models - multimodal embeddings, temporal reasoning, and retrieval built specifically for video rather than adapted from image or text - Real-time and large-scale pipelines - infrastructure for processing, indexing, and querying video at the speed and scale enterprises actually need - Agentic and reasoning layers - systems that don't just retrieve clips but answer questions, surface anomalies, and take action on what they see The models and infrastructure to make this real are appearing to be crossing a capability threshold right now. Multimodal foundation models are maturing, storage costs have collapsed, and enterprises are sitting on years of unstructured video with no way to use it. That infrastructure unlocks a wide range of applications including media and sports workflows, security and physical operations, enterprise knowledge management, advertising analytics, robotics, and consumer products, where video has historically been dark data. If you're building in video intelligence at the model layer, the platform layer, or in a vertical application, we'd love to talk!

Jason Cui

36,205 просмотров • 5 месяцев назад