1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Haoyu Liang, Youran Sun, Yunfeng Cai +2
The security issue of large language models (LLMs) has gained wide attention recently, with various defense mechanisms developed to prevent harmful output, among which safeguards b…