2 papers
cs.LG2025
Primal-Only Actor Critic Algorithm for Robust Constrained Average Cost MDPs
Anirudh Satheesh, Sooraj Sathish, Swetha Ganesh +2
In this work, we study the problem of finding robust and safe policies in Robust Constrained Average-Cost Markov Decision Processes (RCMDPs). A key challenge in this setting is the…
cs.LG2025
Learning Distinguishable Representations in Deep Q-Networks for Linear Transfer
Sooraj Sathish, Keshav Goyal, Raghuram Bharadwaj Diddigi
Deep Reinforcement Learning (RL) has demonstrated success in solving complex sequential decision-making problems by integrating neural networks with the RL framework. However, trai…