1 paper
Richa Verma, Bavish Kulur, Sanjay Chawla +1
We address the problem of making a pre-trained reinforcement learning (RL) policy safety-aware by incorporating cost constraints without retraining it from scratch. While costs cou…