1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.LG2020
Non-asymptotic Convergence of Adam-type Reinforcement Learning Algorithms under Markovian Sampling
Huaqing Xiong, Tengyu Xu, Yingbin Liang +1
Despite the wide applications of Adam in reinforcement learning (RL), the theoretical convergence of Adam-type RL algorithms has not been established. This paper provides the first…
eess.SY2019★ 1 cited
Momentum-based Accelerated Q-learning
Bowen Weng, Lin Zhao, Huaqing Xiong +1
This paper studies accelerated algorithms for Q-learning. We propose an acceleration scheme by incorporating the historical iterates of the Q-function. The idea is conceptually ins…
cs.LG2019
Accelerated Target Updates for Q-learning
Bowen Weng, Huaqing Xiong, Wei Zhang
This paper studies accelerations in Q-learning algorithms. We propose an accelerated target update scheme by incorporating the historical iterates of Q functions. The idea is conce…