1 paper
Daniël Willemsen, Mario Coppola, Guido C. H. E. de Croon
Multi-robot systems can benefit from reinforcement learning (RL) algorithms that learn behaviours in a small number of trials, a property known as sample efficiency. This research…