25 citations · 31 across the 2 of their papers we have counts for
1 paper · 1 filter
Junzi Zhang, Jongho Kim, Brendan O'Donoghue +1
Policy gradient methods are among the most effective methods for large-scale reinforcement learning, and their empirical success has prompted several works that develop the foundat…