activity
20202026
most citedVLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained Models

8 citations · 28 across the 36 of their papers we have counts for

collaborators
Showing cs.CRShow all

10 papers · 1 filter

cs.CR2025

TASO: Jailbreak LLMs via Alternative Template and Suffix Optimization

Yanting Wang, Runpeng Geng, Jinghui Chen +2

Many recent studies showed that LLMs are vulnerable to jailbreak attacks, where an attacker can perturb the input of an LLM to induce it to generate an output for a harmful questio…

cs.CR2025

You Can't Steal Nothing: Mitigating Prompt Leakages in LLMs via System Vectors

Bochuan Cao, Changjiang Li, Yuanpu Cao +3

Large language models (LLMs) have been widely adopted across various applications, leveraging customized system prompts for diverse tasks. Facing potential system prompt leakage ri…

cs.CR2025

Your Agent Can Defend Itself against Backdoor Attacks

Li Changjiang, Liang Jiacheng, Cao Bochuan +2

Despite their growing adoption across domains, large language model (LLM)-powered agents face significant security risks from backdoor attacks during training and fine-tuning. Thes…

cs.CR2025

Towards Robust Multimodal Large Language Models Against Jailbreak Attacks

Ziyi Yin, Yuanpu Cao, Han Liu +3

While multimodal large language models (MLLMs) have achieved remarkable success in recent advancements, their susceptibility to jailbreak attacks has come to light. In such attacks…

cs.CR2024

Data Free Backdoor Attacks

Bochuan Cao, Jinyuan Jia, Chuxuan Hu +5

Backdoor attacks aim to inject a backdoor into a classifier such that it predicts any input with an attacker-chosen backdoor trigger as an attacker-chosen target class. Existing ba…

cs.CR2024★ 1 cited

Watch the Watcher! Backdoor Attacks on Security-Enhancing Diffusion Models

Changjiang Li, Ren Pang, Bochuan Cao +4

Thanks to their remarkable denoising capabilities, diffusion models are increasingly being employed as defensive tools to reinforce the security of other models, notably in purifyi…