1 paper · 1 filter
Kartik Sharma, Yiqiao Jin, Vineeth Rakesh +4
As large language models (LLMs) are deployed in safety-critical settings, it is essential to ensure that their responses comply with safety standards. Prior research has revealed t…