3 papers
cs.LG2025
Initial Distribution Sensitivity of Constrained Markov Decision Processes
Alperen Tercan, Necmiye Ozay
Constrained Markov Decision Processes (CMDPs) are notably more complex to solve than standard MDPs due to the absence of universally optimal policies across all initial state distr…
cs.LG2025
Efficient Reward Identification In Max Entropy Reinforcement Learning with Sparsity and Rank Priors
Mohamad Louai Shehab, Alperen Tercan, Necmiye Ozay
In this paper, we consider the problem of recovering time-varying reward functions from either optimal policies or demonstrations coming from a max entropy reinforcement learning p…
cs.LG2024
Thresholded Lexicographic Ordered Multiobjective Reinforcement Learning
Alperen Tercan, Vinayak S. Prabhu
Lexicographic multi-objective problems, which impose a lexicographic importance order over the objectives, arise in many real-life scenarios. Existing Reinforcement Learning work d…