9 citations · 11 across the 10 of their papers we have counts for
1 paper · 2 filters
Kunhe Yang, Lin F. Yang, Simon S. Du
This paper presents the first non-asymptotic result showing that a model-free algorithm can achieve a logarithmic cumulative regret for episodic tabular reinforcement learning if t…