3 citations · 4 across the 3 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2019★ 1 cited
Transfer Learning Across Simulated Robots With Different Sensors
Hélène Plisnier, Denis Steckelmacher, Diederik Roijers +1
For a robot to learn a good policy, it often requires expensive equipment (such as sophisticated sensors) and a prepared training environment conducive to learning. However, it is…
cs.AI2019★ 3 cited
The Actor-Advisor: Policy Gradient With Off-Policy Advice
Hélène Plisnier, Denis Steckelmacher, Diederik M. Roijers +1
Actor-critic algorithms learn an explicit policy (actor), and an accompanying value function (critic). The actor performs actions in the environment, while the critic evaluates the…