1 paper
Linhao Huang, Xue Jiang, Zhiqiang Wang +5
Video-based multimodal large language models (V-MLLMs) have shown vulnerability to adversarial examples in video-text multimodal tasks. However, the transferability of adversarial…