1 paper
Gianluca Sabatini, Chenhao Li, Marco Hutter
Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively parallel simulation environmen…