1 paper · 1 filter
Osama Hanna, Merve Karakas, Lin F. Yang +1
We consider a novel multi-arm bandit (MAB) setup, where a learner needs to communicate the actions to distributed agents over erasure channels, while the rewards for the actions ar…