8 citations · 23 across the 45 of their papers we have counts for
1 paper · 1 filter
Beitao Chen, Xinyu Lyu, Lianli Gao +2
By incorporating visual inputs, Multimodal Large Language Models (MLLMs) extend LLMs to support visual reasoning. However, this integration also introduces new vulnerabilities, mak…