2 papers
cs.LG2024
Deep reinforcement learning for weakly coupled MDP's with continuous actions
Francisco Robledo, Urtzi Ayesta, Konstantin Avrachenkov
This paper introduces the Lagrange Policy for Continuous Actions (LPCA), a reinforcement learning algorithm specifically designed for weakly coupled MDP problems with continuous ac…
cs.AI2024
Tabular and Deep Learning for the Whittle Index
Francisco Robledo Relaño, Francisco Robledo Relaño, Vivek Borkar +2
The Whittle index policy is a heuristic that has shown remarkably good performance (with guaranteed asymptotic optimality) when applied to the class of problems known as Restless M…