12 citations · 22 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 10 cited
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Xiangxiang Chu, Limeng Qiao, Xinyu Zhang +8
We introduce MobileVLM V2, a family of significantly improved vision language models upon MobileVLM, which proves that a delicate orchestration of novel architectural design, an im…
cs.CV2023★ 12 cited
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Xiangxiang Chu, Limeng Qiao, Xinyang Lin +8
We present MobileVLM, a competent multimodal vision language model (MMVLM) targeted to run on mobile devices. It is an amalgamation of a myriad of architectural designs and techniq…
cs.LG2023
MrTF: Model Refinery for Transductive Federated Learning
Xin-Chun Li, Yang Yang, De-Chuan Zhan
We consider a real-world scenario in which a newly-established pilot project needs to make inferences for newly-collected data with the help of other parties under privacy protecti…