4 papers · 1 filter
Do Input-Level Defenses Transfer to Observation-Level Attacks on VideoLLMs?
Bangshuo Zhu, Wei Song, Yuxin Cao +4
Video Large Language Models (VideoLLMs) are increasingly deployed in safety-critical applications such as content moderation and video analytics. To process long videos efficiently…
VideoSTF: Stress-Testing Output Repetition in Video Large Language Models
Yuxin Cao, Wei Song, Shangzhi Xu +2
Video Large Language Models (VideoLLMs) have recently achieved strong performance in video understanding tasks. However, we identify a previously underexplored generation failure:…
FAIRT2V: Training-Free Debiasing for Text-to-Video Diffusion Models
Haonan Zhong, Wei Song, Tingxu Han +3
Text-to-video (T2V) diffusion models have achieved rapid progress, yet their demographic biases, particularly gender bias, remain largely unexplored. We present FairT2V, a training…
Poisoning Prompt-Guided Sampling in Video Large Language Models
Yuxin Cao, Wei Song, Jingling Xue +1
Video Large Language Models (VideoLLMs) are increasingly deployed as automated moderators on user-generated video platforms, where a few unwatched seconds of harmful footage are en…