3 citations · 3 across the 1 of their papers we have counts for
1 paper
Wangshu Zhu, Andre Rosendo
Proximal policy optimization (PPO) has yielded state-of-the-art results in policy search, a subfield of reinforcement learning, with one of its key points being the use of a surrog…