1 paper
Jiafei Song, Fengwei Zhou, Jin Qu +7
Recent Multimodal Large Language Models (MLLMs) have demonstrated strong performance on vision-language understanding tasks, yet their inference efficiency is often hampered by the…