5 citations · 6 across the 5 of their papers we have counts for
1 paper · 1 filter
Xinwei Zhang, Li Bai, Tianwei Zhang +5
Large vision-language models (LVLMs) have achieved impressive performance across multimodal tasks, but their reliance on visual inputs exposes them to adversarial threats. Encoder-…