2 citations · 2 across the 3 of their papers we have counts for
Showing math.OCShow all
2 papers · 1 filter
math.OC2023
Global Convergence of Policy Gradient Methods in Reinforcement Learning, Games and Control
Shicong Cen, Yuejie Chi
Policy gradient methods, where one searches for the policy of interest by maximizing the value functions using first-order information, become increasingly popular for sequential d…
math.OC2018
A Stochastic Semismooth Newton Method for Nonsmooth Nonconvex Optimization
Andre Milzarek, Xiantao Xiao, Shicong Cen +2
In this work, we present a globalized stochastic semismooth Newton method for solving stochastic optimization problems involving smooth nonconvex and nonsmooth convex terms in the…