67 citations · 165 across the 12 of their papers we have counts for
Showing 2021Show all
2 papers · 1 filter
cs.CV2021★ 67 cited
Image Captioning for Effective Use of Language Models in Knowledge-Based Visual Question Answering
Ander Salaberria, Gorka Azkune, Oier Lopez de Lacalle +2
Integrating outside knowledge for reasoning in visio-linguistic tasks such as visual question answering (VQA) is an open problem. Given that pretrained language models have been sh…
cs.AI2021★ 4 cited
Inferring spatial relations from textual descriptions of images
Aitzol Elu, Gorka Azkune, Oier Lopez de Lacalle +3
Generating an image from its textual description requires both a certain level of language understanding and common sense knowledge about the spatial relations of the physical enti…