2 papers
cs.LG2025
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
Yiting He, Zhishuai Liu, Weixin Wang +1
Off-dynamics reinforcement learning (RL), where training and deployment transition dynamics are different, can be formulated as learning in a robust Markov decision process (RMDP)…
cs.LG2025
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
Jingwen Gu, Yiting He, Zhishuai Liu +1
Decision-making under distribution shift is a central challenge in reinforcement learning (RL), where training and deployment environments differ. We study this problem through the…