1 paper
Michael Muehlebach, Zhiyu He, Michael I. Jordan
We study the sample complexity of online reinforcement learning in the general \hzyrev{non-episodic} setting of nonlinear dynamical systems with continuous state and action spaces.…