1 paper
Qingyun Guo, Junyi Shi, Jianuo Huang +1
Offline reinforcement learning allows control policies to be learned directly from data without online interaction, making it suitable for safety-critical tasks. Recent studies hav…