2 papers
cs.LG2025
GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits
Gongpu Chen, Soung Chang Liew, Deniz Gunduz
The restless multi-armed bandit (RMAB) framework is a popular model with applications across a wide variety of fields. However, its solution is hindered by the exponentially growin…
cs.AI2025
Intermittently Observable Markov Decision Processes
Gongpu Chen, Soung-Chang Liew
This paper investigates MDPs with intermittent state information. We consider a scenario where the controller perceives the state information of the process via an unreliable commu…