18 citations · 18 across the 3 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2023★ 2 cited
Tuning Large language model for End-to-end Speech Translation
Hao Zhang, Nianwen Si, Yaqi Chen +4
With the emergence of large language models (LLMs), multimodal models based on LLMs have demonstrated significant potential. Models such as LLaSM, X-LLM, and SpeechGPT exhibit an i…
cs.CL2023★ 18 cited
Improving Speech Translation by Cross-Modal Multi-Grained Contrastive Learning
Hao Zhang, Nianwen Si, Yaqi Chen +4
The end-to-end speech translation (E2E-ST) model has gradually become a mainstream paradigm due to its low latency and less error propagation. However, it is non-trivial to train s…
cs.CL2023
Decouple Non-parametric Knowledge Distillation For End-to-end Speech Translation
Hao Zhang, Nianwen Si, Yaqi Chen +4
Existing techniques often attempt to make knowledge transfer from a powerful machine translation (MT) to speech translation (ST) model with some elaborate techniques, which often r…