1 paper
David Martínez-Rubio, Varun Kanade, Patrick Rebeschini
We study a decentralized cooperative stochastic multi-armed bandit problem with K arms on a network of N agents. In our model, the reward distribution of each arm is the same f…