1 paper
Caglar Gulcehre, Srivatsan Srinivasan, Jakub Sygnowski +5
Deep neural networks are the most commonly used function approximators in offline reinforcement learning. Prior works have shown that neural nets trained with TD-learning and gradi…