7 citations · 22 across the 11 of their papers we have counts for
1 paper · 1 filter
Letitia Parcalabescu, Albert Gatt, Anette Frank +1
We investigate the reasoning ability of pretrained vision and language (V&L) models in two tasks that require multimodal integration: (1) discriminating a correct image-sentence pa…