42 citations · 89 across the 23 of their papers we have counts for
34 papers · 1 filter
Incremental Blockwise Beam Search for Simultaneous Speech Translation with Controllable Quality-Latency Tradeoff
Peter Polák, Brian Yan, Shinji Watanabe +2
Blockwise self-attentional encoder models have recently emerged as one promising end-to-end approach to simultaneous speech translation. These models employ a blockwise beam search…
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
Christian Huber, Tu Anh Dinh, Carlos Mullov +10
The challenge of low-latency speech translation has recently draw significant interest in the research community as shown by several publications and shared tasks. Therefore, it is…
KIT's Multilingual Speech Translation System for IWSLT 2023
Danni Liu, Thai Binh Nguyen, Sai Koneru +7
Many existing speech translation benchmarks focus on native-English speech in high-quality recording conditions, which often do not match the conditions in real-life use-cases. In…
Code-Switching without Switching: Language Agnostic End-to-End Speech Translation
Christian Huber, Enes Yavuz Ugan, Alexander Waibel
We propose a) a Language Agnostic end-to-end Speech Translation model (LAST), and b) a data augmentation strategy to increase code-switching (CS) performance. With increasing globa…
Adaptive multilingual speech recognition with pretrained models
Ngoc-Quan Pham, Alex Waibel, Jan Niehues
Multilingual speech recognition with supervised learning has achieved great results as reflected in recent research. With the development of pretraining methods on audio and text d…
CUNI-KIT System for Simultaneous Speech Translation Task at IWSLT 2022
Peter Polák, Ngoc-Quan Ngoc, Tuan-Nam Nguyen +5
In this paper, we describe our submission to the Simultaneous Speech Translation at IWSLT 2022. We explore strategies to utilize an offline model in a simultaneous setting without…