10 citations · 19 across the 11 of their papers we have counts for
Showing 2024 · cs.LGShow all
3 papers · 2 filters
cs.LG2024★ 1 cited
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
Huizhen Yu, Yi Wan, Richard S. Sutton
This paper investigates the stability and convergence properties of asynchronous stochastic approximation (SA) algorithms, with a focus on extensions relevant to average-reward rei…
cs.LG2024
On Convergence of Average-Reward Q-Learning in Weakly Communicating Markov Decision Processes
Yi Wan, Huizhen Yu, Richard S. Sutton
This paper analyzes reinforcement learning (RL) algorithms for Markov decision processes (MDPs) under the average-reward criterion. We focus on Q-learning algorithms based on relat…
cs.LG2024★ 3 cited
Reward Centering
Abhishek Naik, Yi Wan, Manan Tomar +1
We show that discounted methods for solving continuing reinforcement learning problems can perform significantly better if they center their rewards by subtracting out the rewards'…