1 paper · 1 filter
Weidi Luo, He Cao, Zijing Liu +5
With the extensive deployment of Large Language Models (LLMs), ensuring their safety has become increasingly critical. However, existing defense methods often struggle with two key…