3 papers
math.OC2025
Stochastic Primal-Dual Q-Learning
Narim Jeong, Donghwan Lee, Niao He
In this work, we present a new model-free and off-policy reinforcement learning (RL) algorithm, that is capable of finding a near-optimal policy with state-action observations from…
math.OC2025
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
Donghwan Lee
In this paper, the primary goal is to offer additional insights into the value iteration through the lens of switching system models in the control community. These models establis…
cs.LG2025
Regularized Q-learning
Han-Dong Lim, Donghwan Lee
Q-learning is widely used algorithm in reinforcement learning community. Under the lookup table setting, its convergence is well established. However, its behavior is known to be u…