1 paper
Jingyuan Liu, Hao Qiu, Lin Yang +1
We study the distributed multi-agent multi-armed bandit problem with heterogeneous rewards over random communication graphs. Uniquely, at each time step t agents communicate over…