2 citations · 7 across the 9 of their papers we have counts for
1 paper · 2 filters
Wenliang Zhong, Wenyi Wu, Qi Li +6
Multimodal Large Language Models (MLLMs) have achieved SOTA performance in various visual language tasks by fusing the visual representations with LLMs leveraging some visual adapt…