2 papers
math.OC2025
Mirror descent for constrained stochastic control problems
Deven Sethi, David Šiška
Mirror descent is a well established tool for solving convex optimization problems with convex constraints. This article introduces continuous-time mirror descent dynamics for appr…
math.OC2025
Entropy annealing for policy mirror descent in continuous time and space
Deven Sethi, David Šiška, Yufei Zhang
Entropy regularization has been widely used in policy optimization algorithms to enhance exploration and the robustness of the optimal control; however it also introduces an additi…