1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Lu Guo, Yixiang Shan, Zhengbang Zhu +5
Offline reinforcement learning (RL) learns policies from fixed datasets, thereby avoiding costly or unsafe environment interactions. However, its reliance on finite static datasets…