21 citations · 21 across the 2 of their papers we have counts for
1 paper · 1 filter
Yimin Shi
Deep Reinforcement Learning (DRL) sometimes needs a large amount of data to converge in the training procedure and in some cases, each action of the agent may produce regret. This…