1 paper
Zhaorun Chen, Zhuokai Zhao, Tairan He +4
Ensuring safety in Reinforcement Learning (RL), typically framed as a Constrained Markov Decision Process (CMDP), is crucial for real-world exploration applications. Current approa…