2 citations · 2 across the 9 of their papers we have counts for
1 paper · 1 filter
Casimir Czworkowski, Stephen Hornish, Alhassan S. Yasin
Proximal Policy Optimization (PPO) is a widely used reinforcement learning algorithm known for its stability and sample efficiency, but it often suffers from premature convergence…