1 paper
Hao Liu, Yihao Feng, Yi Mao +3
Policy gradient methods have achieved remarkable successes in solving challenging reinforcement learning problems. However, it still often suffers from the large variance issue on…