1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
Vishesh Mittal, Rahul Meshram, Surya Prakash
We study the Whittle index learning algorithm for restless multi-armed bandits. We consider index learning algorithm with Q-learning. We first present Q-learning algorithm with exp…
cs.LG2023★ 1 cited
Indexability of Finite State Restless Multi-Armed Bandit and Rollout Policy
Vishesh Mittal, Rahul Meshram, Deepak Dev +1
We consider finite state restless multi-armed bandit problem. The decision maker can act on M bandits out of N bandits in each time step. The play of arm (active arm) yields state…