1 citations · 1 across the 3 of their papers we have counts for
3 papers
Faster Q-Learning Algorithms for Restless Bandits
Parvish Kakarapalli, Devendra Kayande, Rahul Meshram
We study the Whittle index learning algorithm for restless multi-armed bandits (RMAB). We first present Q-learning algorithm and its variants -- speedy Q-learning (SQL), generalize…
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
Vishesh Mittal, Rahul Meshram, Surya Prakash
We study the Whittle index learning algorithm for restless multi-armed bandits. We consider index learning algorithm with Q-learning. We first present Q-learning algorithm with exp…
Indexability of Finite State Restless Multi-Armed Bandit and Rollout Policy
Vishesh Mittal, Rahul Meshram, Deepak Dev +1
We consider finite state restless multi-armed bandit problem. The decision maker can act on M bandits out of N bandits in each time step. The play of arm (active arm) yields state…