1 paper · 1 filter
Pulkit Katdare, Anant Joshi, Katherine Driggs-Campbell
Policy gradient methods are a vital ingredient behind the success of modern reinforcement learning. Modern policy gradient methods, although successful, introduce a residual error…