1 paper · 1 filter
Yuanzhe Shen, Zisu Huang, Zhengkang Guo +5
The rapid advancement of large language models (LLMs) has driven their adoption across diverse domains, yet their ability to generate harmful content poses significant safety chall…