5 citations · 5 across the 1 of their papers we have counts for
1 paper
Xiaoyu Wen, Xudong Yu, Rui Yang +3
To obtain a near-optimal policy with fewer interactions in Reinforcement Learning (RL), a promising approach involves the combination of offline RL, which enhances sample efficienc…