9 citations · 9 across the 9 of their papers we have counts for
1 paper · 1 filter
Kunhe Yang, Lin F. Yang, Simon S. Du
This paper presents the first non-asymptotic result showing that a model-free algorithm can achieve a logarithmic cumulative regret for episodic tabular reinforcement learning if t…