4 citations · 9 across the 12 of their papers we have counts for
1 paper · 2 filters
Xu Ji, Jianyi Zhang, Ziyin Zhou +5
Ensuring the resilience of Large Language Models (LLMs) against malicious exploitation is paramount, with recent focus on mitigating offensive responses. Yet, the understanding of…