1 paper · 1 filter
Yuyao Sun, Tao Deng, Shuang Li +3
Multimodal large language models (MLLMs) achieve strong performance across diverse vision-language tasks, but their efficiency is limited by the cost of processing numerous visual…