1 paper
Liyuan Mao, Haoran Xu, Weinan Zhang +1
In this study, we investigate the DIstribution Correction Estimation (DICE) methods, an important line of work in offline reinforcement learning (RL) and imitation learning (IL). D…