4 papers
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
Renyang Liu, Guanlin Li, Tianwei Zhang +1
Recent advances in image generation models (IGMs), particularly diffusion-based architectures such as Stable Diffusion (SD), have markedly enhanced the quality and diversity of AI-…
Picky LLMs and Unreliable RMs: An Empirical Study on Safety Alignment after Instruction Tuning
Guanlin Li, Kangjie Chen, Shangwei Guo +6
Large language models (LLMs) have emerged as powerful tools for addressing a wide range of general inquiries and tasks. Despite this, fine-tuning aligned LLMs on smaller, domain-sp…
Warfare:Breaking the Watermark Protection of AI-Generated Content
Guanlin Li, Yifei Chen, Jie Zhang +5
AI-Generated Content (AIGC) is rapidly expanding, with services using advanced generative models to create realistic images and fluent text. Regulating such content is crucial to p…
ART: Automatic Red-teaming for Text-to-Image Models to Protect Benign Users
Guanlin Li, Kangjie Chen, Shudong Zhang +2
Large-scale pre-trained generative models are taking the world by storm, due to their abilities in generating creative content. Meanwhile, safeguards for these generative models ar…