43 citations · 63 across the 4 of their papers we have counts for
1 paper · 1 filter
Long Yang, Jiaming Ji, Juntao Dai +5
Safe reinforcement learning (RL) studies problems where an intelligent agent has to not only maximize reward but also avoid exploring unsafe areas. In this study, we propose CUP, a…