8 citations · 8 across the 1 of their papers we have counts for
2 papers
cs.LG2020
Finite-Time Analysis for Double Q-learning
Huaqing Xiong, Lin Zhao, Yingbin Liang +1
Although Q-learning is one of the most successful algorithms for finding the best action-value function (and thus the optimal policy) in reinforcement learning, its implementation…
cs.LG2020★ 8 cited
Momentum Q-learning with Finite-Sample Convergence Guarantee
Bowen Weng, Huaqing Xiong, Lin Zhao +2
Existing studies indicate that momentum ideas in conventional optimization can be used to improve the performance of Q-learning algorithms. However, the finite-sample analysis for…