1 paper
Hannah Sterz, Jonas Pfeiffer, Ivan VuliÄ
Vision Language Models (VLMs) extend remarkable capabilities of text-only large language models and vision-only models, and are able to learn from and process multi-modal vision-te…