4 citations · 7 across the 13 of their papers we have counts for
1 paper · 1 filter
Yihe Deng, Yu Yang, Junkai Zhang +2
The rapid advancement of large language models (LLMs) necessitates effective mechanisms to ensure their responsible deployment by accurately distinguishing unsafe content from beni…