1 paper
Jiamin Xu, Ivan Nazarov, Aditya Rastogi +2
This paper addresses the poor finite-horizon performance of existing online \emph{restless bandit} (RB) algorithms, which stems from the prohibitive sample complexity of learning a…