6 citations · 6 across the 6 of their papers we have counts for
1 paper · 1 filter
Zhepeng Cen, Zuxin Liu, Zitong Wang +3
Offline reinforcement learning (RL) offers a promising direction for learning policies from pre-collected datasets without requiring further interactions with the environment. Howe…