1 paper
Wenhui Zhu, Xin Li, Xiwen Chen +8
Recently, Multimodal Large Language Models (MLLMs) have gained significant attention for their remarkable ability to process and analyze non-textual data, such as images, videos, a…