Video yükleniyor...
Video Yüklenemedi
Reinforcement Learning for Robotics Part 2: Train a Balance Bot with PPO with Shawn Hymel Full video:
14,017 görüntüleme • 3 gün önce •via X (Twitter)
1 Yorum

Mandar Wagh2 gün önce
@ShawnHymel the number that decides whether it transfers: a 30 cm inverted pendulum has a time constant of sqrt(L/g), about 175 ms, and the error grows by a factor of e over that. sim has zero latency, so every millisecond of real sensor and actuator delay comes out of that budget.
