1 paper · 1 filter
Xiaohan Hu, Yi Ma, Chenjun Xiao +2
One of the fundamental challenges for offline reinforcement learning (RL) is ensuring robustness to data distribution. Whether the data originates from a near-optimal policy or not…