1 paper
Arash Bahari Kordabad, Dean Brandner, Sebastien Gros +2
In this paper, we propose a second-order deterministic actor-critic framework in reinforcement learning that extends the classical deterministic policy gradient method to exploit c…