1 paper · 1 filter
Haolin Liu, Braham Snyder, Chen-Yu Wei
We study offline reinforcement learning under Q⋆-approximation and partial coverage, a setting that motivates practical algorithms such as Conservative Q-Learning (CQL; Ku…