1 paper · 1 filter
Sanghyun Son, Laura Yu Zheng, Ryan Sullivan +2
We introduce a novel policy learning method that integrates analytical gradients from differentiable environments with the Proximal Policy Optimization (PPO) algorithm. To incorpor…