1 paper
Seungyub Han, Hyungjin Kim, Jungwoo Lee
Offline reinforcement learning (RL) agents often fail when deployed, as the gap between training datasets and real environments leads to unsafe behavior. To address this, we presen…