1 paper
Paul Mattes, Rainer Schlosser, Ralf Herbrich
One of the biggest challenges to modern deep reinforcement learning (DRL) algorithms is sample efficiency. Many approaches learn a world model in order to train an agent entirely i…