3 papers
stat.ML2025
Accelerated Distributional Temporal Difference Learning with Linear Function Approximation
Kaicheng Jin, Yang Peng, Jiansheng Yang +1
In this paper, we study the finite-sample statistical rates of distributional temporal difference (TD) learning with linear function approximation. The purpose of distributional TD…
stat.ML2025
A Finite Sample Analysis of Distributional TD Learning with Linear Function Approximation
Yang Peng, Kaicheng Jin, Liangyu Zhang +1
In this paper, we study the finite-sample statistical rates of distributional temporal difference (TD) learning with linear function approximation. The aim of distributional TD lea…
math.OC2025
A Regularized Online Newton Method for Stochastic Convex Bandits with Linear Vanishing Noise
Jingxin Zhan, Yuchen Xin, Kaicheng Jin +1
We study a stochastic convex bandit problem where the subgaussian noise parameter is assumed to decrease linearly as the learner selects actions closer and closer to the minimizer…