1 paper
Chongming Gao, Kexin Huang, Jiawei Chen +6
Offline reinforcement learning (RL), a technology that offline learns a policy from logged data without the need to interact with online environments, has become a favorable choice…