12 citations · 22 across the 3 of their papers we have counts for
1 paper · 1 filter
Rui Yang, Han Zhong, Jiawei Xu +4
Offline reinforcement learning (RL) presents a promising approach for learning reinforced policies from offline datasets without the need for costly or unsafe interactions with the…