1 paper
Minjae Kwon, Josephine Lamp, Lu Feng
Safe Reinforcement Learning (RL) algorithms are typically evaluated under fixed training conditions. We investigate whether training-time safety guarantees transfer to deployment u…