4 citations · 5 across the 6 of their papers we have counts for
1 paper · 2 filters
Gongwei Chen, Leyang Shen, Rui Shao +2
Multimodal Large Language Models (MLLMs) have endowed LLMs with the ability to perceive and understand multi-modal signals. However, most of the existing MLLMs mainly adopt vision…