1 paper · 1 filter
Yixiu Mao, Qi Wang, Chen Chen +2
In offline reinforcement learning (RL), addressing the out-of-distribution (OOD) action issue has been a focus, but we argue that there exists an OOD state issue that also impairs…