6 citations · 6 across the 1 of their papers we have counts for
3 papers
eess.AS2020★ 6 cited
DenoiSpeech: Denoising Text to Speech with Frame-Level Noise Modeling
Chen Zhang, Yi Ren, Xu Tan +5
While neural-based text to speech (TTS) models can synthesize natural and intelligible voice, they usually require high-quality speech data, which is costly to collect. In many sce…
eess.AS2020
FastLR: Non-Autoregressive Lipreading Model with Integrate-and-Fire
Jinglin Liu, Yi Ren, Zhou Zhao +3
Lipreading is an impressive technique and there has been a definite improvement of accuracy in recent years. However, existing methods for lipreading mainly build on autoregressive…
eess.AS2020
UWSpeech: Speech to Speech Translation for Unwritten Languages
Chen Zhang, Xu Tan, Yi Ren +3
Existing speech to speech translation systems heavily rely on the text of target language: they usually translate source language either to target text and then synthesize target s…