1 paper
Charles Blundell, Benigno Uria, Alexander Pritzel +6
State of the art deep reinforcement learning algorithms take many millions of interactions to attain human-level performance. Humans, on the other hand, can very quickly exploit hi…