1 paper · 1 filter
Bardia Azizian, Ivan V. Bajic
The rapid progress of large Vision-Language Models (VLMs) has enabled a wide range of applications, such as image understanding and Visual Question Answering (VQA). Query images ar…