Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Accelerating Suffix Jailbreak attacks with Prefix-Shared KV-cache
Xinhai Wang, Shaopeng Fu, Shu Yang +3
Suffix jailbreak attacks serve as a systematic method for red-teaming Large Language Models (LLMs) but suffer from prohibitive computational costs, as a large number of candidate s…
cs.CR2025
Backdooring CLIP through Concept Confusion
Lijie Hu, Junchi Liao, Weimin Lyu +5
Backdoor attacks pose a serious threat to deep learning models by allowing adversaries to implant hidden behaviors that remain dormant on clean inputs but are maliciously triggered…