1 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Tianyu Sun
We develop Policy Gradient with Second-Order Momentum (PG-SOM), a lightweight second-order optimisation scheme for reinforcement-learning policies. PG-SOM augments the classical RE…