7 citations · 41 across the 23 of their papers we have counts for
1 paper · 1 filter
Prathamesh Mayekar, Jonathan Scarlett, Vincent Y. F. Tan
We study a distributed stochastic multi-armed bandit where a client supplies the learner with communication-constrained feedback based on the rewards for the corresponding arm pull…