1 paper
Xinda Qi, Dong Chen, Zhaojian Li +1
In this paper, we propose a novel technique, Back-stepping Experience Replay (BER), that is compatible with arbitrary off-policy reinforcement learning (RL) algorithms. BER aims to…