11 citations · 11 across the 1 of their papers we have counts for
1 paper
Eric Chen, Zhang-Wei Hong, Joni Pajarinen +1
State-of-the-art reinforcement learning (RL) algorithms typically use random sampling (e.g., ε-greedy) for exploration, but this method fails on hard exploration tasks like Monte…