9 citations · 19 across the 8 of their papers we have counts for
4 papers · 1 filter
SRTNet: Time Domain Speech Enhancement Via Stochastic Refinement
Zhibin Qiu, Mengfan Fu, Yinfeng Yu +3
Diffusion model, as a new generative model which is very popular in image generation and audio synthesis, is rarely used in speech enhancement. In this paper, we use the diffusion…
Leveraging Phone Mask Training for Phonetic-Reduction-Robust E2E Uyghur Speech Recognition
Guodong Ma, Pengfei Hu, Jian Kang +2
In Uyghur speech, consonant and vowel reduction are often encountered, especially in spontaneous speech with high speech rate, which will cause a degradation of speech recognition…
Enriching Under-Represented Named-Entities To Improve Speech Recognition Performance
Tingzhi Mao, Yerbolat Khassanov, Van Tung Pham +4
Automatic speech recognition (ASR) for under-represented named-entity (UR-NE) is challenging due to such named-entities (NE) have insufficient instances and poor contextual coverag…
Mandarin tone modeling using recurrent neural networks
Hao Huang, Ying Hu, Haihua Xu
We propose an Encoder-Classifier framework to model the Mandarin tones using recurrent neural networks (RNN). In this framework, extracted frames of features for tone classificatio…