Showing 2023 · math.OCShow all
2 papers · 2 filters
math.OC2023
On the Convergence of Projected Policy Gradient for Any Constant Step Sizes
Jiacai Liu, Wenye Li, Dachao Lin +2
Projected policy gradient (PPG) is a basic policy optimization method in reinforcement learning. Given access to exact policy evaluations, previous studies have established the sub…
math.OC2023
On the Linear Convergence of Policy Gradient under Hadamard Parameterization
Jiacai Liu, Jinchi Chen, Ke Wei
The convergence of deterministic policy gradient under the Hadamard parameterization is studied in the tabular setting and the linear convergence of the algorithm is established. T…