30 citations · 43 across the 3 of their papers we have counts for
1 paper · 1 filter
Ming Yu, Zhuoran Yang, Mladen Kolar +1
We study the safe reinforcement learning problem with nonlinear function approximation, where policy optimization is formulated as a constrained optimization problem with both the…