2 papers
cs.CL2025
Jailbreaking Commercial Black-Box LLMs with Explicitly Harmful Prompts
Chiyu Zhang, Lu Zhou, Xiaogang Xu +3
Existing black-box jailbreak attacks achieve certain success on non-reasoning models but degrade significantly on recent SOTA reasoning models. To improve attack ability, inspired…
cs.CV2024
Adversarial Attacks of Vision Tasks in the Past 10 Years: A Survey
Chiyu Zhang, Lu Zhou, Xiaogang Xu +2
With the advent of Large Vision-Language Models (LVLMs), new attack vectors, such as cognitive bias, prompt injection, and jailbreaking, have emerged. Understanding these attacks p…