1 paper · 1 filter
Xu Ji, Jianyi Zhang, Ziyin Zhou +5
Ensuring the resilience of Large Language Models (LLMs) against malicious exploitation is paramount, with recent focus on mitigating offensive responses. Yet, the understanding of…