1 paper · 1 filter
Hongqiang Lin, Zhenghui Fu, Weihao Tang +4
Offline reinforcement learning (RL) enables data-efficient and safe policy learning without online exploration, but its performance often degrades under distribution shift. The lea…