14 citations · 15 across the 4 of their papers we have counts for
1 paper · 2 filters
Yang Zheng, Yue Sun, Maryam Fazel +1
First order policy optimization has been widely used in reinforcement learning. It guarantees to find the optimal policy for the state-feedback linear quadratic regulator (LQR). Ho…