1 paper
Zhuoyan Xu, Khoi Duc Nguyen, Preeti Mukherjee +4
Multimodal Large Language Models (MLLMs) have shown impressive capabilities in visual reasoning, yet come with substantial computational cost, limiting their deployment in resource…