1 paper · 1 filter
Yuling Yan, Gen Li, Yuxin Chen +1
This paper is concerned with the asynchronous form of Q-learning, which applies a stochastic approximation scheme to Markovian data samples. Motivated by the recent advances in off…