1 paper
Yulin Chen, Haoran Li, Yirui Zhang +3
Multimodal Large Language Models (MLLMs) have showcased impressive performance in a variety of multimodal tasks. On the other hand, the integration of additional image modality may…