Loading video...
Video Failed to Load
Model-Free Reinforcement Learning (MFRL) has been alluring, especially with supercharged compute with physics on GPU. However, the methods use 0-th order gradients, and are often not the best optimizers. Can we do better than PPO in continuous control for robotics? Turns out yes! 🥳 tl;dr: Faster, better RL than... show more
52,308 views • 2 years ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
