753 citations · 1.1k across the 3 of their papers we have counts for
1 paper · 1 filter
William Fedus, Prajit Ramachandran, Rishabh Agarwal +4
Experience replay is central to off-policy algorithms in deep reinforcement learning (RL), but there remain significant gaps in our understanding. We therefore present a systematic…