1 paper
Conor Newton, Ayalvadi Ganesh, Henry W. J. Reeve
We consider a large number of agents collaborating on a multi-armed bandit problem with a large number of arms. The goal is to minimise the regret of each agent in a communication-…