2 papers
cs.LG2025
Improving Transferability of Adversarial Examples via Bayesian Attacks
Qizhang Li, Yiwen Guo, Xiaochen Yang +2
The transferability of adversarial examples allows for the attack on unknown deep neural networks (DNNs), posing a serious threat to many applications and attracting great attentio…
cs.LG2025
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
Qizhang Li, Xiaochen Yang, Wangmeng Zuo +1
Automatic adversarial prompt generation provides remarkable success in jailbreaking safely-aligned large language models (LLMs). Existing gradient-based attacks, while demonstratin…