1 citations · 1 across the 1 of their papers we have counts for
5 papers
Cascade versus Direct Speech Translation: Do the Differences Still Make a Difference?
Luisa Bentivogli, Mauro Cettolo, Marco Gaido +4
Five years after the first published proofs of concept, direct approaches to speech translation (ST) are now competing with traditional cascade solutions. In light of this steady p…
Beyond Voice Activity Detection: Hybrid Audio Segmentation for Direct Speech Translation
Marco Gaido, Matteo Negri, Mauro Cettolo +1
The audio segmentation mismatch between training data and those seen at run-time is a major problem in direct speech translation. Indeed, while systems are usually trained on manua…
CTC-based Compression for Direct Speech Translation
Marco Gaido, Mauro Cettolo, Matteo Negri +1
Previous studies demonstrated that a dynamic phone-informed compression of the input audio is beneficial for speech translation (ST). However, they required a dedicated model for p…
Contextualized Translation of Automatically Segmented Speech
Marco Gaido, Mattia Antonino Di Gangi, Matteo Negri +2
Direct speech-to-text translation (ST) models are usually trained on corpora segmented at sentence level, but at inference time they are commonly fed with audio split by a voice ac…
A Comparison of Transformer and Recurrent Neural Networks on Multilingual Neural Machine Translation
Surafel M. Lakew, Mauro Cettolo, Marcello Federico
Recently, neural machine translation (NMT) has been extended to multilinguality, that is to handle more than one translation direction with a single system. Multilingual NMT showed…