1 paper
Zhishuai Liu, Weixin Wang, Pan Xu
We study off-dynamics Reinforcement Learning (RL), where the policy training and deployment environments are different. To deal with this environmental perturbation, we focus on le…