1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Hongfu Liu, Hengguan Huang, Xiangming Gu +2
Large language models (LLMs) pose significant risks due to the potential for generating harmful content or users attempting to evade guardrails. Existing studies have developed LLM…