1 paper · 1 filter
Ayush Rai, Shaoshuai Mou
Multi-armed bandit algorithms provide solutions for sequential decision-making where learning takes place by interacting with the environment. In this work, we model a distributed…