3 citations · 4 across the 5 of their papers we have counts for
1 paper · 1 filter
Wangshu Zhu, Andre Rosendo
Proximal policy optimization (PPO) has yielded state-of-the-art results in policy search, a subfield of reinforcement learning, with one of its key points being the use of a surrog…