134 citations · 134 across the 1 of their papers we have counts for
1 paper
Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen +5
Most deep reinforcement learning (RL) systems are not able to learn effectively from off-policy data, especially if they cannot explore online in the environment. These are critica…