2 papers
cs.CR2026
Why Are LLM Backdoor Defenses Fragmented? A Feature-Level Explanation with Sparse Autoencoders
Yizhe Zeng, Chenxu Niu, Wei Zhang +7
Backdoor attacks pose a serious threat to large language models (LLMs), but existing defenses remain fragmented, failing to pro?vide unified defense against both dirty-label and cl…
cs.CV2026
BadDreamer: Transferable Backdoor Attacks against Video World Models for Autonomous Driving
Zhe Shuai, Xiaopeng Xie, Yikun Zeng
Video world models are increasingly used in autonomous driving to forecast future scene evolution and provide future-aware spatio-temporal representations for downstream action pre…