3 citations · 6 across the 8 of their papers we have counts for
15 papers
MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages
Marco Gaido, Sara Papi, Luisa Bentivogli +6
The rise of foundation models (FMs), coupled with regulatory efforts addressing their risks and impacts, has sparked significant interest in open-source models. However, existing s…
Who Are We Talking About? Handling Person Names in Speech Translation
Marco Gaido, Matteo Negri, Marco Turchi
Recent work has shown that systems for speech translation (ST) -- similarly to automatic speech recognition (ASR) -- poorly handle person names. This shortcoming does not only lead…
Under the Morphosyntactic Lens: A Multifaceted Evaluation of Gender Bias in Speech Translation
Beatrice Savoldi, Marco Gaido, Luisa Bentivogli +2
Gender bias is largely recognized as a problematic phenomenon affecting language technologies, with recent studies underscoring that it might surface differently across languages.…
Is "moby dick" a Whale or a Bird? Named Entities and Terminology in Speech Translation
Marco Gaido, Susana Rodríguez, Matteo Negri +2
Automatic translation systems are known to struggle with rare words. Among these, named entities (NEs) and domain-specific terms are crucial, since errors in their translation can…
Between Flexibility and Consistency: Joint Generation of Captions and Subtitles
Alina Karakanta, Marco Gaido, Matteo Negri +1
Speech translation (ST) has lately received growing interest for the generation of subtitles without the need for an intermediate source language transcription and timing (i.e. cap…
Cascade versus Direct Speech Translation: Do the Differences Still Make a Difference?
Luisa Bentivogli, Mauro Cettolo, Marco Gaido +4
Five years after the first published proofs of concept, direct approaches to speech translation (ST) are now competing with traditional cascade solutions. In light of this steady p…