58 citations · 115 across the 13 of their papers we have counts for
1 paper · 2 filters
Chao Yu, Akash Velu, Eugene Vinitsky +4
Proximal Policy Optimization (PPO) is a ubiquitous on-policy reinforcement learning algorithm but is significantly less utilized than off-policy learning algorithms in multi-agent…