1 paper · 1 filter
Michael Orme, Yanchao Yu, Zhiyuan Tan
Ensuring safe and contextually appropriate behaviour in Large Language Models (LLMs) remains a critical challenge for real-world deployment. We present \textbf{SafeCtrl-RL}, an inf…