2 papers
cs.LG2019
Investigation on the generalization of the Sampled Policy Gradient algorithm
Nil Stolt Ansó
The Sampled Policy Gradient (SPG) algorithm is a new offline actor-critic variant that samples in the action space to approximate the policy gradient. It does so by using the criti…
cs.AI2018
Sampled Policy Gradient for Learning to Play the Game Agar.io
Anton Orell Wiehe, Nil Stolt Ansó, Madalina M. Drugan +1
In this paper, a new offline actor-critic learning algorithm is introduced: Sampled Policy Gradient (SPG). SPG samples in the action space to calculate an approximated policy gradi…