1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Parvish Kakarapalli, Devendra Kayande, Rahul Meshram
We study the Whittle index learning algorithm for restless multi-armed bandits (RMAB). We first present Q-learning algorithm and its variants -- speedy Q-learning (SQL), generalize…