1 paper
Yuyao Sun, Tao Deng, Shuang Li +3
Multimodal large language models (MLLMs) achieve strong performance across diverse vision-language tasks, but their efficiency is limited by the cost of processing numerous visual…