34 citations · 54 across the 3 of their papers we have counts for
1 paper · 1 filter
Yue Wu, Shuangfei Zhai, Nitish Srivastava +4
Offline Reinforcement Learning promises to learn effective policies from previously-collected, static datasets without the need for exploration. However, existing Q-learning and ac…