1 paper
Weiqin Chen, Xinjie Zhang, Sandipan Mishra +1
Offline reinforcement learning (RL) learns effective policies from a static target dataset. The performance of state-of-the-art offline RL algorithms notwithstanding, it relies on…