正在加载视频...
视频加载失败
Reinforcement Learning for Robotics Part 2: Train a Balance Bot with PPO with Shawn Hymel Full video:
14,017 次观看 • 3 天前 •via X (Twitter)
1 条评论

Mandar Wagh2 天前
@ShawnHymel the number that decides whether it transfers: a 30 cm inverted pendulum has a time constant of sqrt(L/g), about 175 ms, and the error grows by a factor of e over that. sim has zero latency, so every millisecond of real sensor and actuator delay comes out of that budget.
