1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Gongwei Chen, Leyang Shen, Rui Shao +2
Multimodal Large Language Models (MLLMs) have endowed LLMs with the ability to perceive and understand multi-modal signals. However, most of the existing MLLMs mainly adopt vision…