2 papers
cs.LG2026
Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations
Jun Liu, Leo Yu Zhang, Fengpeng Li +2
Hard-label black-box attacks, relying solely on top-1 predictions, represent one of the most challenging yet practically threat models. Despite recent progress, existing approaches…
cs.LG2026
AEGIS: Adversarial Target-Guided Retention-Data-Free Robust Concept Erasure from Diffusion Models
Fengpeng Li, Kemou Li, Qizhou Wang +2
Concept erasure helps stop diffusion models (DMs) from generating harmful content; but current methods face robustness retention trade off. Robustness means the model fine-tuned by…