2 citations · 2 across the 3 of their papers we have counts for
4 papers
Bootstrapping Audiovisual Speech Recognition in Zero-AV-Resource Scenarios with Synthetic Visual Data
Pol Buitrago, Pol Gàlvez, Oriol Pareras +1
Audiovisual speech recognition (AVSR) combines acoustic and visual cues to improve transcription robustness under challenging conditions but remains out of reach for most under-res…
Quantifying Cross-Lingual Transfer in Paralinguistic Speech Tasks
Pol Buitrago, Oriol Pareras, Federico Costa +1
Paralinguistic speech tasks are often considered relatively language-agnostic, as they rely on extralinguistic acoustic cues rather than lexical content. However, prior studies rep…
Breaking Language Barriers in Visual Language Models via Multilingual Textual Regularization
Iñigo Pikabea, Iñaki Lacunza, Oriol Pareras +4
Rapid advancements in Visual Language Models (VLMs) have transformed multimodal understanding but are often constrained by generating English responses regardless of the input lang…
Salamandra Technical Report
Aitor Gonzalez-Agirre, Marc Pàmies, Joan Llop +21
This work introduces Salamandra, a suite of open-source decoder-only large language models available in three different sizes: 2, 7, and 40 billion parameters. The models were trai…