activity
20242026
collaborators

41 papers

cs.CV2026

CIVA: Critic-Induced Value-Subspace Attacks on Visual World-Model Agents

Jiancheng Wang, Mingli Zhu, Tong Zhang +4

Visual world-model agents such as DreamerV3 act through a recurrent latent state rather than a single observation, which weakens frame-wise observation attacks and makes their pert…

cs.CV2026

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense

Siyuan Liang, Yupeng Qiu, Junfeng Fang +3

Text-to-Video (T2V) generative models are vulnerable to jailbreak attacks in real-world deployment, leading them to produce harmful or inappropriate content. Existing defense appro…

cs.CR2026

Breadcrumbing Search Agents

Xuebin Li, Hanqing Zhao, Siyuan Liang +4

LLM-based search agents are widely used for information-seeking tasks, but their reliance on external tool returns introduces a critical security risk: web content retrieved during…

cs.CR2026

SoK: Intent-Oriented Systematization of Multi-Turn LLM Jailbreaks

Siyuan Li, Aodu Wulianghai, Zehao Liu +8

Large Language Models (LLMs) are increasingly deployed in interactive settings, where user intent commonly unfolds through multi-turn dialogue. Multi-turn jailbreaks exploit this p…

cs.CV2026

Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective

Tianyuan Zhang, Xianglong Liu, Aishan Liu +6

Environmental illusions (eg., shadows, reflections, and tire marks) are naturally existing yet overlooked phenomena in real-world driving environments. They can disturb visual perc…

cs.CR2026

Cert-LAS: Toward Certified Model Ownership Verification for Text-to-Image Diffusion Models via Layer-Adaptive Smoothing

Leyi Qi, Yiming Li, Siyuan Liang +2

Large-scale text-to-image (T2I) diffusion models have enabled unprecedented creative applications, but their unauthorized use has raised serious intellectual property concerns, mak…