2 papers
math.OC2026
Discretization error from regularized Reinforcement Learning to continuous-time stochastic control
Huyên Pham, Yuming Paul Zhang, Yuhua Zhu
This paper establishes a rigorous connection between regularized discrete-time reinforcement learning (RL) and continuous-time stochastic optimal control. Specifically, classical R…
math.OC2025
Policy iteration for the deterministic control problems -- a viscosity approach
Wenpin Tang, Hung Vinh Tran, Yuming Paul Zhang
This paper is concerned with the convergence rate of policy iteration for (deterministic) optimal control problems in continuous time. To overcome the problem of ill-posedness due…