1 paper
Jean-Charles Layoun, Alexis Roger, Irina Rish
The goal of vision-language modeling is to allow models to tie language understanding with visual inputs. The aim of this paper is to evaluate and align the Visual Language Model (…