1 citations · 1 across the 2 of their papers we have counts for
Showing math.OCShow all
2 papers · 1 filter
math.OC2026
Discretization error from regularized Reinforcement Learning to continuous-time stochastic control
Huyên Pham, Yuming Paul Zhang, Yuhua Zhu
This paper establishes a rigorous connection between regularized discrete-time reinforcement learning (RL) and continuous-time stochastic optimal control. Specifically, classical R…
math.OC2021★ 1 cited
Exploratory HJB equations and their convergence
Wenpin Tang, Paul Yuming Zhang, Xun Yu Zhou
We study the exploratory Hamilton--Jacobi--Bellman (HJB) equation arising from the entropy-regularized exploratory control problem, which was formulated by Wang, Zariphopoulou and…