1 citations · 1 across the 1 of their papers we have counts for
1 paper
Guanghe Li, Yixiang Shan, Zhengbang Zhu +2
In offline reinforcement learning (RL), the performance of the learned policy highly depends on the quality of offline datasets. However, in many cases, the offline dataset contain…