activity
20172026
most citedOver-Generation Cannot Be Rewarded: Length-Adaptive Average Lagging for Simultaneous Speech Translation

18 citations · 92 across the 29 of their papers we have counts for

collaborators
Showing 2022 · cs.CLShow all

9 papers · 2 filters

cs.CL2022★ 7 cited

Attention as a Guide for Simultaneous Speech Translation

Sara Papi, Matteo Negri, Marco Turchi

The study of the attention mechanism has sparked interest in many fields, such as language modeling and machine translation. Although its patterns have been exploited to perform di…

cs.CL2022★ 5 cited

Joint Speech Translation and Named Entity Recognition

Marco Gaido, Sara Papi, Matteo Negri +1

Modern automatic translation systems aim at place the human at the center by providing contextual support and knowledge. In this context, a critical task is enriching the output wi…

cs.CL2022

Direct Speech Translation for Automatic Subtitling

Sara Papi, Marco Gaido, Alina Karakanta +3

Automatic subtitling is the task of automatically translating the speech of audiovisual content into short pieces of timed text, i.e. subtitles and their corresponding timestamps.…

cs.CL2022★ 2 cited

Dodging the Data Bottleneck: Automatic Subtitling with Automatically Segmented ST Corpora

Sara Papi, Alina Karakanta, Matteo Negri +1

Speech translation for subtitling (SubST) is the task of automatically translating speech data into well-formed subtitles by inserting subtitle breaks compliant to specific display…

cs.CL2022★ 18 cited

Over-Generation Cannot Be Rewarded: Length-Adaptive Average Lagging for Simultaneous Speech Translation

Sara Papi, Marco Gaido, Matteo Negri +1

Simultaneous speech translation (SimulST) systems aim at generating their output with the lowest possible latency, which is normally computed in terms of Average Lagging (AL). In t…

cs.CL2022

Who Are We Talking About? Handling Person Names in Speech Translation

Marco Gaido, Matteo Negri, Marco Turchi

Recent work has shown that systems for speech translation (ST) -- similarly to automatic speech recognition (ASR) -- poorly handle person names. This shortcoming does not only lead…