7 citations · 14 across the 4 of their papers we have counts for
1 paper · 1 filter
Tom Blau, Lionel Ott, Fabio Ramos
Balancing exploration and exploitation is a fundamental part of reinforcement learning, yet most state-of-the-art algorithms use a naive exploration protocol like ε-greedy. This…