1 paper
Yingen Liu, Fan Wu, Ruihui Li +2
Multimodal large language models (MLLMs) demonstrate strong performance across visual tasks, but their efficiency is hindered by significant computational and memory demands from p…