1 paper
Ming Yin, Mengdi Wang, Yu-Xiang Wang
This article reviews the recent advances on the statistical foundation of reinforcement learning (RL) in the offline and low-adaptive settings. We will start by arguing why offline…