3 citations · 6 across the 7 of their papers we have counts for
5 papers · 1 filter
Multimodal Large Language Models for Low-Resource Languages: A Case Study for Basque
Lukas Arana, Julen Etxaniz, Ander Salaberria +1
Current Multimodal Large Language Models exhibit very strong performance for several demanding tasks. While commercial MLLMs deliver acceptable performance in low-resource language…
Vision-Language Models Struggle to Align Entities across Modalities
Iñigo Alonso, Gorka Azkune, Ander Salaberria +2
Cross-modal entity linking refers to the ability to align entities and their attributes across different modalities. While cross-modal entity linking is a fundamental skill needed…
Grounding Spatial Relations in Text-Only Language Models
Gorka Azkune, Ander Salaberria, Eneko Agirre
This paper shows that text-only Language Models (LM) can learn to ground spatial relations like "left of" or "below" if they are provided with explicit location information of obje…
IXA/Cogcomp at SemEval-2023 Task 2: Context-enriched Multilingual Named Entity Recognition using Knowledge Bases
Iker García-Ferrero, Jon Ander Campos, Oscar Sainz +2
Named Entity Recognition (NER) is a core natural language processing task in which pre-trained language models have shown remarkable performance. However, standard benchmarks like…
Evaluating Multimodal Representations on Visual Semantic Textual Similarity
Oier Lopez de Lacalle, Ander Salaberria, Aitor Soroa +2
The combination of visual and textual representations has produced excellent results in tasks such as image captioning and visual question answering, but the inference capabilities…