1 paper
Zhaoqi Xu, Yingying Zhang, Jian Li +3
Recent advances in vision-language models (VLMs) have shown remarkable performance across multimodal tasks, yet their ever-growing scale poses severe challenges for deployment and…