10 papers
SpreadMark: Robust Image Watermarking via Spread-Spectrum Embedding
Wei Song, Yuxin Cao, Zhenchang Xing +4
Invisible image watermarks are increasingly used for deepfake detection and provenance tracking, where they must survive not only incidental distortions but also deliberate removal…
Caved or Convinced: Temporal Sampling Gates Claim Deference in Video Large Language Models
Yuxin Cao, Wei Song, Jingling Xue +1
When asked which of two events came first, video large language models can fail in two opposite ways: cave to a false claim, or reject a true one. Prior video sycophancy work measu…
Membership Inference Attacks Against Video Large Language Models
Wei Song, Yuxin Cao, Ziqi Ding +3
Video large language models (VideoLLMs) are increasingly trained or instruction-tuned on large-scale video--text corpora collected from heterogeneous sources, raising an immediate…
DUAP: Dual-task Universal Adversarial Perturbations Against Voice Control Systems
Suyang Sun, Weifei Jin, Yuxin Cao +2
Modern Voice Control Systems (VCS) rely on the collaboration of Automatic Speech Recognition (ASR) and Speaker Recognition (SR) for secure interaction. However, prior adversarial a…
VideoSTF: Stress-Testing Output Repetition in Video Large Language Models
Yuxin Cao, Wei Song, Shangzhi Xu +2
Video Large Language Models (VideoLLMs) have recently achieved strong performance in video understanding tasks. However, we identify a previously underexplored generation failure:…
FAIRT2V: Training-Free Debiasing for Text-to-Video Diffusion Models
Haonan Zhong, Wei Song, Tingxu Han +3
Text-to-video (T2V) diffusion models have achieved rapid progress, yet their demographic biases, particularly gender bias, remain largely unexplored. We present FairT2V, a training…