1 paper
Jing Fu, Bill Moran, Jose Nino-Mora
We study a finite time horizon Markov decision process (MDP) consisting of several groups of multi-action finite-state restless bandit processes, which are identical within each gr…