Showing math.OCShow all
3 papers · 1 filter
math.OC2025
Stochastic Primal-Dual Q-Learning
Narim Jeong, Donghwan Lee, Niao He
In this work, we present a new model-free and off-policy reinforcement learning (RL) algorithm, that is capable of finding a near-optimal policy with state-action observations from…
math.OC2025
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
Donghwan Lee
In this paper, the primary goal is to offer additional insights into the value iteration through the lens of switching system models in the control community. These models establis…
math.OC2024
Lossless Convexification and Duality
Donghwan Lee
The main goal of this paper is to investigate strong duality of non-convex semidefinite programming problems (SDPs). In the optimization community, it is well-known that a convex o…