1 paper
Weichao Li, Fuxian Huang, Xi Li +2
A critical and challenging problem in reinforcement learning is how to learn the state-action value function from the experience replay buffer and simultaneously keep sample effici…