1 paper
Qian He, Wenqi Liang, Chunhui Hao +2
Mimicking the real interaction trajectory in the inference of the world model has been shown to improve the sample efficiency of model-based reinforcement learning (MBRL) algorithm…