正在加载视频...

视频加载失败

Using reinforcement learning, we trained policies for Boston Dynamics Spot that allow the robot to achieve record running speeds of 11.5 mph (5.2 m/s) — over three times faster than Spot's default max speed.

98,538 次观看 • 1 年前 •via X (Twitter)

11 条评论

Loic Argelies 的头像
Loic Argelies1 年前

@BostonDynamics Need a gallop mode…

The Rundown AI 的头像
The Rundown AI1 年前

If you're not learning AI in 2025, you're falling behind. Join 1,000,000+ early adopters reading and learn AI in just 5 minutes a day (for free).

io.net 的头像
io.net1 年前

@BostonDynamics At 11.5 mph, Spot is officially faster than most humans. Sleep tight.

Victor Brink 的头像
Victor Brink1 年前

@BostonDynamics Great! Now teach it how a change in gait like a cheetah can multiply its top speed further... 110 km/hr (70 mph) This would be epic. The Black Mirror episode featuring a running bot is "Metalhead" for further inspiration.

Elie Aljalbout 的头像
Elie Aljalbout1 年前

@BostonDynamics very impressive!

TeslaElon SpaceXFan 的头像
TeslaElon SpaceXFan1 年前

@BostonDynamics 🥳👍

Liberty Mint 的头像
Liberty Mint1 年前

@BostonDynamics For sustained duration? Faster speeds achieved in sprint?

AI_TechnoKing 的头像
AI_TechnoKing1 年前

@BostonDynamics Galloping gears!

Rodomonte 的头像
Rodomonte1 年前

@BostonDynamics @AskPerplexity give me calculations on max possible physical speed for a robot of that weight

Pat Dunne 的头像
Pat Dunne1 年前

@BostonDynamics Yup RL.. just like motors.. it’s a thing .. welcome BD back to the race

Tom Ramirez 的头像
Tom Ramirez1 年前

@BostonDynamics Robot PT

相关视频

Today, we're joined by Nikita Rudin, co-founder and CEO of Flexion to discuss the gap between current robotic capabilities and what’s required to deploy fully autonomous robots in the real world. Nikita explains how reinforcement learning and simulation have driven rapid progress in robot locomotion—and why locomotion is still far from “solved.” We dig into the sim2real gap, and how adding visual inputs introduces noise and significantly complicates sim-to-real transfer. We also explore the debate between end-to-end models and modular approaches, and why separating locomotion, planning, and semantics remains a pragmatic approach today. Nikita also introduces the concept of "real-to-sim", which uses real-world data to refine simulation parameters for higher fidelity training, discusses how reinforcement learning, imitation learning, and teleoperation data are combined to train robust policies for both quadruped and humanoid robots, and introduces Flexion's hierarchical approach that utilizes pre-trained Vision-Language Models (VLMs) for high-level task orchestration with Vision-Language-Action (VLA) models and low-level whole-body trackers. Finally, Nikita shares the behind-the-scenes in humanoid robot demos, his take on reinforcement learning in simulation versus the real world, the nuances of reward tuning, and offers practical advice for researchers and practitioners looking to get started in robotics today. 🗒️ For the full list of resources for this episode, visit the show notes page: 📖 CHAPTERS =============================== 00:00 - Introduction 04:07 - Is robot locomotion solved? 06:04 - Sim-to-real gap 08:58 - Adding semantics to policies 09:42 - Modular vs end-to-end architectures 10:29 - Planner model 12:21 - Adapting RL techniques from quadrupeds to humanoids 15:39 - Behind robot demos 18:09 - Humanoid robots in home environments 22:03 - Training approach 23:56 - VLA models 27:59 - Closing the sim-to-real gap 32:55 - Task orchestration using VLMs 36:38 - Tool use 38:10 - Model hierarchy 43:37 - Simulator versus simulation environment 44:57 - Combining imitation learning and reinforcement learning 46:42 - RL in real world versus RL in simulation 52:58 - Reward tuning and value functions in robotics 56:38 - Predictions 1:00:10 - Humanoids, quadropeds, and wheeled platforms 1:02:45 - Advice, recommended robot kits, and community pla

The TWIML AI Podcast

22,582 次观看 • 6 个月前