1 paper
Ravi Shekhar, Sandro Pezzelle, Yauhen Klimovich +4
In this paper, we aim to understand whether current language and vision (LaVi) models truly grasp the interaction between the two modalities. To this end, we propose an extension o…