3 papers
cs.CR2026
From ASR to ASP: Evaluating Prompt Attack Vulnerabilities Against Open-Source LLMs
Jiawen Wang, Pritha Gupta, Ivan Habernal +3
Recent studies demonstrate that Large Language Models (LLMs) are vulnerable to attacks that generate harmful or sensitive outputs. As open-source LLMs are increasingly adopted in h…
cs.CR2025
Practical, Generalizable and Robust Backdoor Attacks on Text-to-Image Diffusion Models
Haoran Dai, Jiawen Wang, Ruo Yang +4
Text-to-image diffusion models (T2I DMs) have achieved remarkable success in generating high-quality and diverse images from text prompts, yet recent studies have revealed their vu…
cs.CR2025
Backdoor Attacks on Discrete Graph Diffusion Models
Jiawen Wang, Samin Karim, Yuan Hong +1
Diffusion models are powerful generative models in continuous data domains such as image and video data. Discrete graph diffusion models (DGDMs) have recently extended them for gra…