1 paper
Xi Chen, Ali Ghadirzadeh, Tianhe Yu +6
Offline reinforcement learning methods hold the promise of learning policies from pre-collected datasets without the need to query the environment for new transitions. This setting…