1 paper
Shaowei Guan, Yu Zhai, Zhengyu Zhang +2
Large Language Models (LLMs) are increasingly vulnerable to adversarial attacks that can subtly manipulate their outputs. While various defense mechanisms have been proposed, many…