3 papers
cs.LG2026
Collaborating in Multi-Armed Bandits with Strategic Agents
Idan Barnea, Ofir Schlisselberg, Yishay Mansour
We study collaborative learning in multi-agent Bayesian bandit problems, where strategic agents collectively solve the same bandit instance. While multiple agents can accelerate le…
cs.LG2026
The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration
Idan Barnea, Orin Levy, Yishay Mansour
We study cooperative multi-agent reinforcement learning in the setting of reward-free exploration, where multiple agents jointly explore an unknown MDP in order to learn its dynami…
cs.LG2026
Individual Regret in Cooperative Stochastic Multi-Armed Bandits
Idan Barnea, Tal Lancewicki, Yishay Mansour
We study the regret in stochastic Multi-Armed Bandits (MAB) with multiple agents that communicate over an arbitrary connected communication graph. We analyzed a variant of Cooperat…