1 paper
Shilin Xu, Xiangtai Li, Haobo Yuan +3
The recent surge in Multimodal Large Language Models (MLLMs) has showcased their remarkable potential for achieving generalized intelligence by integrating visual understanding int…