activity
20182021
most citedCascade versus Direct Speech Translation: Do the Differences Still Make a Difference?

1 citations · 1 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL20211 cited

Cascade versus Direct Speech Translation: Do the Differences Still Make a Difference?

Luisa Bentivogli, Mauro Cettolo, Marco Gaido +4

Five years after the first published proofs of concept, direct approaches to speech translation (ST) are now competing with traditional cascade solutions. In light of this steady p…

cs.SD2021

Beyond Voice Activity Detection: Hybrid Audio Segmentation for Direct Speech Translation

Marco Gaido, Matteo Negri, Mauro Cettolo +1

The audio segmentation mismatch between training data and those seen at run-time is a major problem in direct speech translation. Indeed, while systems are usually trained on manua…

cs.CL2021

CTC-based Compression for Direct Speech Translation

Marco Gaido, Mauro Cettolo, Matteo Negri +1

Previous studies demonstrated that a dynamic phone-informed compression of the input audio is beneficial for speech translation (ST). However, they required a dedicated model for p…

cs.CL2020

Contextualized Translation of Automatically Segmented Speech

Marco Gaido, Mattia Antonino Di Gangi, Matteo Negri +2

Direct speech-to-text translation (ST) models are usually trained on corpora segmented at sentence level, but at inference time they are commonly fed with audio split by a voice ac…

cs.CL2018

A Comparison of Transformer and Recurrent Neural Networks on Multilingual Neural Machine Translation

Surafel M. Lakew, Mauro Cettolo, Marcello Federico

Recently, neural machine translation (NMT) has been extended to multilinguality, that is to handle more than one translation direction with a single system. Multilingual NMT showed…