1 paper
Xuehui Yu, Yi Guan, Rujia Shen +3
Model-based offline Reinforcement Learning (RL) allows agents to fully utilise pre-collected datasets without requiring additional or unethical explorations. However, applying mode…