1 paper
Poojitha Thota, Yun Lei, Santhosh Thangaraj +2
Large language models (LLMs) are increasingly deployed in interactive applications, yet they remain vulnerable to adversarial interactions that induce harmful, deceptive, or policy…