2 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CL2023
Improving speech translation by fusing speech and text
Wenbiao Yin, Zhicheng Liu, Chengqi Zhao +3
In speech translation, leveraging multimodal data to improve model performance and address limitations of individual modalities has shown significant effectiveness. In this paper,…
cs.CL2023★ 2 cited
DUB: Discrete Unit Back-translation for Speech Translation
Dong Zhang, Rong Ye, Tom Ko +2
How can speech-to-text translation (ST) perform as well as machine translation (MT)? The key point is to bridge the modality gap between speech and text so that useful MT technique…
cs.CL2022★ 2 cited
On the Impact of Noises in Crowd-Sourced Data for Speech Translation
Siqi Ouyang, Rong Ye, Lei Li
Training speech translation (ST) models requires large and high-quality datasets. MuST-C is one of the most widely used ST benchmark datasets. It contains around 400 hours of speec…