1 paper
Pusen Dong, Tianchen Zhu, Yue Qiu +2
Safe reinforcement learning (RL) requires the agent to finish a given task while obeying specific constraints. Giving constraints in natural language form has great potential for p…