4 papers
Localize and Neutralize: Gradient-guided Token Suppression against Visual Prompt Injection Attack
Dongpeng Zhang, Ke Ma, Yangbangyan Jiang +4
Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principled understanding of the unde…
Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain
Gaozheng Pei, Ke Ma, Yingfei Sun +2
The diffusion-based adversarial purification methods attempt to drown adversarial perturbations into a part of isotropic noise through the forward process, and then recover the cle…
A Unified Framework for Stealthy Adversarial Generation via Latent Optimization and Transferability Enhancement
Gaozheng Pei, Ke Ma, Dongpeng Zhang +3
Due to their powerful image generation capabilities, diffusion-based adversarial example generation methods through image editing are rapidly gaining popularity. However, due to re…
Exploring Query Efficient Data Generation towards Data-free Model Stealing in Hard Label Setting
Gaozheng Pei, Shaojie lyu, Ke Ma +3
Data-free model stealing involves replicating the functionality of a target model into a substitute model without accessing the target model's structure, parameters, or training da…