25 citations · 25 across the 1 of their papers we have counts for
1 paper
Rongjun Qin, Songyi Gao, Xingyuan Zhang +5
Offline reinforcement learning (RL) aims at learning a good policy from a batch of collected data, without extra interactions with the environment during training. However, current…