6 citations · 19 across the 7 of their papers we have counts for
1 paper · 1 filter
Kaiqing Zhang, Bin Hu, Tamer Başar
Policy optimization (PO) is a key ingredient for reinforcement learning (RL). For control design, certain constraints are usually enforced on the policies to optimize, accounting f…