4 citations · 4 across the 6 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
Rahul Meshram, Kesav Kaza
Partially observable restless multi-armed bandits have found numerous applications including in recommendation systems, communication systems, public healthcare outreach systems, a…
cs.LG2021
Indexability and Rollout Policy for Multi-State Partially Observable Restless Bandits
Rahul Meshram, Kesav Kaza
Restless multi-armed bandits with partially observable states has applications in communication systems, age of information and recommendation systems. In this paper, we study mult…