1 paper
Hyeju Shin, Chorwon Kim, Ryangsoo Kim +2
The emergence of vision language models with fewer than 3 billion parameters has accelerated the implementation of on-device multimodal intelligence. However, a detailed understand…