1 paper · 1 filter
Ivo Danihelka
Many reinforcement learning exploration techniques are overly optimistic and try to explore every state. Such exploration is impossible in environments with the unlimited number of…