1 paper
Ehsan Imani, Eric Graves, Martha White
Policy gradient methods are widely used for control in reinforcement learning, particularly for the continuous action setting. There have been a host of theoretically sound algorit…