2 papers
cs.LG2025
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
Jingwen Gu, Yiting He, Zhishuai Liu +1
Decision-making under distribution shift is a central challenge in reinforcement learning (RL), where training and deployment environments differ. We study this problem through the…
cs.LG2024
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
Zhishuai Liu, Weixin Wang, Pan Xu
We study off-dynamics Reinforcement Learning (RL), where the policy training and deployment environments are different. To deal with this environmental perturbation, we focus on le…