1 paper
Yaqi Duan, Martin J. Wainwright
We introduce a novel framework for analyzing reinforcement learning (RL) in continuous state-action spaces, and use it to prove fast rates of convergence in both off-line and on-li…