5 citations · 5 across the 1 of their papers we have counts for
1 paper
Xingxing Liang, Yang Ma, Yanghe Feng +1
On-policy deep reinforcement learning algorithms have low data utilization and require significant experience for policy improvement. This paper proposes a proximal policy optimiza…