8 citations · 34 across the 13 of their papers we have counts for
4 papers · 2 filters
Diformer: Directional Transformer for Neural Machine Translation
Minghan Wang, Jiaxin Guo, Yuxia Wang +8
Autoregressive (AR) and Non-autoregressive (NAR) models have their own superiority on the performance and latency, combining them into one model may take advantage of both. Current…
Joint-training on Symbiosis Networks for Deep Nueral Machine Translation models
Zhengzhe Yu, Jiaxin Guo, Minghan Wang +11
Deep encoders have been proven to be effective in improving neural machine translation (NMT) systems, but it reaches the upper bound of translation quality when the number of encod…
Self-Distillation Mixup Training for Non-autoregressive Neural Machine Translation
Jiaxin Guo, Minghan Wang, Daimeng Wei +11
Recently, non-autoregressive (NAT) models predict outputs in parallel, achieving substantial improvements in generation speed compared to autoregressive (AT) models. While performi…
The HW-TSC's Offline Speech Translation Systems for IWSLT 2021 Evaluation
Minghan Wang, Yuxia Wang, Chang Su +9
This paper describes our work in participation of the IWSLT-2021 offline speech translation task. Our system was built in a cascade form, including a speaker diarization module, an…