1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Andreas Doerr, Michael Volpp, Marc Toussaint +2
Policy gradient methods are powerful reinforcement learning algorithms and have been demonstrated to solve many complex tasks. However, these methods are also data-inefficient, aff…