1 paper
Shaikat Galib, Shanshan Wang, Guanshuo Xu +4
Smaller vision-language models (VLMs) are becoming increasingly important for privacy-focused, on-device applications due to their ability to run efficiently on consumer hardware f…