117 citations · 184 across the 11 of their papers we have counts for
1 paper · 2 filters
Jarryd Martin, Suraj Narayanan Sasikumar, Tom Everitt +1
We introduce a new count-based optimistic exploration algorithm for Reinforcement Learning (RL) that is feasible in environments with high-dimensional state-action spaces. The succ…