808 citations · 1.7k across the 5 of their papers we have counts for
1 paper · 1 filter
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki +4
Deep reinforcement learning agents have achieved state-of-the-art results by directly maximising cumulative reward. However, environments contain a much wider variety of possible t…