1 citations · 2 across the 17 of their papers we have counts for
1 paper · 1 filter
Byung-Kwan Lee, Yu-Chiang Frank Wang, Ryo Hachiuma
Large-scale vision-language models (VLMs) have recently achieved remarkable multimodal understanding, but their massive size makes them impractical for deployment on mobile or edge…