22 citations · 24 across the 4 of their papers we have counts for
5 papers
Cross-modal Contrastive Learning for Speech Translation
Rong Ye, Mingxuan Wang, Lei Li
How can we learn unified representations for spoken utterances and their written text? Learning similar representations for semantically similar speech and text is important for sp…
STEMM: Self-learning with Speech-text Manifold Mixup for Speech Translation
Qingkai Fang, Rong Ye, Lei Li +2
How to learn a better speech representation for end-to-end speech-to-text translation (ST) with limited labeled data? Existing techniques often attempt to transfer powerful machine…
The Volctrans Neural Speech Translation System for IWSLT 2021
Chengqi Zhao, Zhicheng Liu, Jian Tong +6
This paper describes the systems submitted to IWSLT 2021 by the Volctrans team. We participate in the offline speech translation and text-to-text simultaneous translation tracks. F…
End-to-end Speech Translation via Cross-modal Progressive Training
Rong Ye, Mingxuan Wang, Lei Li
End-to-end speech translation models have become a new trend in research due to their potential of reducing error propagation. However, these models still suffer from the challenge…
Variational Template Machine for Data-to-Text Generation
Rong Ye, Wenxian Shi, Hao Zhou +2
How to generate descriptions from structured data organized in tables? Existing approaches using neural encoder-decoder models often suffer from lacking diversity. We claim that an…