1 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Jianren Wang, Yifan Su, Abhinav Gupta +1
On-policy reinforcement learning (RL) algorithms are widely used for their strong asymptotic performance and training stability, but they struggle to scale with larger batch sizes,…