2 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Ziqi Zhao, Zhaochun Ren, Liu Yang +6
Offline reinforcement learning (RL) aims to learn policies without online explorations. To enlarge the training data, model-based offline RL learns a dynamics model which is utiliz…