1 paper
Francisco Robledo, Urtzi Ayesta, Konstantin Avrachenkov
This paper introduces the Lagrange Policy for Continuous Actions (LPCA), a reinforcement learning algorithm specifically designed for weakly coupled MDP problems with continuous ac…