2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Yanjie Li, Lina Yu, Weijun Li +6
Current multimodal large language models (MLLMs) are mainly focused on the understanding and processing of perceptual modalities such as images and videos, while their capability f…