12 citations · 27 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 10 cited
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Xiangxiang Chu, Limeng Qiao, Xinyu Zhang +8
We introduce MobileVLM V2, a family of significantly improved vision language models upon MobileVLM, which proves that a delicate orchestration of novel architectural design, an im…
cs.CV2023★ 12 cited
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Xiangxiang Chu, Limeng Qiao, Xinyang Lin +8
We present MobileVLM, a competent multimodal vision language model (MMVLM) targeted to run on mobile devices. It is an amalgamation of a myriad of architectural designs and techniq…
cs.CV2023★ 5 cited
PivotNet: Vectorized Pivot Learning for End-to-end HD Map Construction
Wenjie Ding, Limeng Qiao, Xi Qiu +1
Vectorized high-definition map online construction has garnered considerable attention in the field of autonomous driving research. Most existing approaches model changeable map el…