4 papers
IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization
Zixuan Chen, Jiaxiang Chen, Li Luo +4
LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on untrusted external content expo…
Dynamic Optimization and Safety Indicator Injection for Jailbreaking Text-to-Image Models with Multimodal Safety Filters
Zixuan Chen, Hao Lin, Ke Xu +2
Text-to-image (T2I) models can generate not-safe-for-work (NSFW) content, motivating multi-stage safety pipelines with both text and image filters. Newer LLM-based filters detect l…
FGAS: Fixed Decoder Network-Based Audio Steganography with Adversarial Perturbation Generation
Jialin Yan, Yu Cheng, Zhaoxia Yin +4
The rapid development of Artificial Intelligence Generated Content (AIGC) has made high-fidelity generated audio widely available across the Internet, driving the advancement of au…
AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction
Zixuan Chen, Depeng Wang, Hao Lin +6
We present AVID, the first large-scale benchmark for audio-visual inconsistency understanding in videos. While omni-modal large language models excel at temporally aligned tasks su…