1 paper
Andreas Koukounas, Georgios Mastrapas, Florian Hönicke +5
We present jina-vlm, a token-efficient 2.4B parameter vision-language model that achieves state-of-the-art multilingual VQA performance among open 2B-scale VLMs. The model couples…