32 citations · 32 across the 1 of their papers we have counts for
2 papers
cs.LG2018
Finite Sample Analysis of LSTD with Random Projections and Eligibility Traces
Haifang Li, Yingce Xia, Wensheng Zhang
Policy evaluation with linear function approximation is an important problem in reinforcement learning. When facing high-dimensional feature spaces, such a problem becomes extremel…
cs.LG2015★ 32 cited
Thompson Sampling for Budgeted Multi-armed Bandits
Yingce Xia, Haifang Li, Tao Qin +2
Thompson sampling is one of the earliest randomized algorithms for multi-armed bandits (MAB). In this paper, we extend the Thompson sampling to Budgeted MAB, where there is random…