127 citations · 330 across the 23 of their papers we have counts for
1 paper · 2 filters
Samarth Sinha, Jiaming Song, Animesh Garg +1
The use of past experiences to accelerate temporal difference (TD) learning of value functions, or experience replay, is a key component in deep reinforcement learning. Prioritizat…