collaborators

9 papers

cs.CR2026

SpreadMark: Robust Image Watermarking via Spread-Spectrum Embedding

Wei Song, Yuxin Cao, Zhenchang Xing +4

Invisible image watermarks are increasingly used for deepfake detection and provenance tracking, where they must survive not only incidental distortions but also deliberate removal…

cs.MM2026

Caved or Convinced: Temporal Sampling Gates Claim Deference in Video Large Language Models

Yuxin Cao, Wei Song, Jingling Xue +1

When asked which of two events came first, video large language models can fail in two opposite ways: cave to a false claim, or reject a true one. Prior video sycophancy work measu…

cs.CV2026

VideoSTF: Stress-Testing Output Repetition in Video Large Language Models

Yuxin Cao, Wei Song, Shangzhi Xu +2

Video Large Language Models (VideoLLMs) have recently achieved strong performance in video understanding tasks. However, we identify a previously underexplored generation failure:…

cs.CR2026

DeMark: A Query-Free Black-Box Attack on Deepfake Watermarking Defenses

Wei Song, Zhenchang Xing, Liming Zhu +2

The rapid proliferation of realistic deepfakes has raised urgent concerns over their misuse, motivating the use of defensive watermarks in synthetic images for reliable detection a…

cs.SD2026

Robust CAPTCHA Using Audio Illusions in the Era of Large Language Models: from Evaluation to Advances

Ziqi Ding, Yunfeng Wan, Wei Song +7

CAPTCHAs are widely used by websites to block bots and spam by presenting challenges that are easy for humans but difficult for automated programs to solve. To improve accessibilit…

cs.MM2025

Failures to Surface Harmful Contents in Video Large Language Models

Yuxin Cao, Wei Song, Derui Wang +2

Video Large Language Models (VideoLLMs) are increasingly deployed on numerous critical applications, where users rely on auto-generated summaries while casually skimming the video…