4 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Junming Yang, Xingguo Chen, Shengyuan Wang +1
Model-based offline reinforcement learning (RL), which builds a supervised transition model with logging dataset to avoid costly interactions with the online environment, has been…