2 papers
cs.AI2024
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
Krishna C. Kalagarla, Dhruva Kartik, Dongming Shen +3
Autonomous systems often have logical constraints arising, for example, from safety, operational, or regulatory requirements. Such constraints can be expressed using temporal logic…
cs.LG2024
ACPO: A Policy Optimization Algorithm for Average MDPs with Constraints
Akhil Agnihotri, Rahul Jain, Haipeng Luo
Reinforcement Learning (RL) for constrained MDPs (CMDPs) is an increasingly important problem for various applications. Often, the average criterion is more suitable than the disco…