42 citations · 47 across the 4 of their papers we have counts for
1 paper · 1 filter
Yihuan Mao, Chao Wang, Bin Wang +1
With the success of offline reinforcement learning (RL), offline trained RL policies have the potential to be further improved when deployed online. A smooth transfer of the policy…