2 papers
cs.LG2024
Faster Q-Learning Algorithms for Restless Bandits
Parvish Kakarapalli, Devendra Kayande, Rahul Meshram
We study the Whittle index learning algorithm for restless multi-armed bandits (RMAB). We first present Q-learning algorithm and its variants -- speedy Q-learning (SQL), generalize…
cs.LG2024
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
Vishesh Mittal, Rahul Meshram, Surya Prakash
We study the Whittle index learning algorithm for restless multi-armed bandits. We consider index learning algorithm with Q-learning. We first present Q-learning algorithm with exp…