1 paper · 1 filter
Merve Karakas, Osama Hanna, Lin F. Yang +1
In this paper, we consider a multi-armed bandit (MAB) instance and study how to identify the best arm when arm commands are conveyed from a central learner to a distributed agent o…