19 citations · 19 across the 1 of their papers we have counts for
1 paper
Chang Tian, An Liu, Guang Huang +1
We propose a successive convex approximation based off-policy optimization (SCAOPO) algorithm to solve the general constrained reinforcement learning problem, which is formulated a…