2 papers
cs.LG2025
Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features
Zixuan Xie, Xinyu Liu, Rohan Chandra +1
Linear TD() is one of the most fundamental reinforcement learning algorithms for policy evaluation. Previously, convergence rates are typically established under the assumption…
cs.LG2025
Linear -Learning Does Not Diverge in : Convergence Rates to a Bounded Set
Xinyu Liu, Zixuan Xie, Shangtong Zhang
-learning is one of the most fundamental reinforcement learning algorithms. It is widely believed that -learning with linear function approximation (i.e., linear -learning…