1 paper · 1 filter
David Brellmann, Eloïse Berthier, David Filliat +1
Temporal Difference (TD) algorithms are widely used in Deep Reinforcement Learning (RL). Their performance is heavily influenced by the size of the neural network. While in supervi…