1 paper
Valentin Charvet, Sebastian Stein, Roderick Murray-Smith
Controllers trained with Reinforcement Learning tend to be very specialized and thus generalize poorly when their testing environment differs from their training one. We propose a…