7 citations · 7 across the 1 of their papers we have counts for
1 paper · 1 filter
Long Yang, Jiaming Ji, Juntao Dai +3
Safe reinforcement learning (RL) is still very challenging since it requires the agent to consider both return maximization and safe exploration. In this paper, we propose CUP, a C…