1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2020
A Decentralized Policy with Logarithmic Regret for a Class of Multi-Agent Multi-Armed Bandit Problems with Option Unavailability Constraints and Stochastic Communication Protocols
Pathmanathan Pankayaraj, D. H. S. Maithripala, J. M. Berg
This paper considers a multi-armed bandit (MAB) problem in which multiple mobile agents receive rewards by sampling from a collection of spatially dispersed stochastic processes, c…
cs.LG2019★ 1 cited
A Decentralized Communication Policy for Multi Agent Multi Armed Bandit Problems
Pathmanathan Pankayaraj, D. H. S. Maithripala
This paper proposes a novel policy for a group of agents to, individually as well as collectively, solve a multi armed bandit (MAB) problem. The policy relies solely on the informa…