1 citations · 1 across the 1 of their papers we have counts for
1 paper
Gongwei Chen, Leyang Shen, Rui Shao +2
Multimodal Large Language Models (MLLMs) have endowed LLMs with the ability to perceive and understand multi-modal signals. However, most of the existing MLLMs mainly adopt vision…