12 citations · 12 across the 2 of their papers we have counts for
2 papers
eess.AS2023
Spike-Triggered Contextual Biasing for End-to-End Mandarin Speech Recognition
Kaixun Huang, Ao Zhang, Binbin Zhang +3
The attention-based deep contextual biasing method has been demonstrated to effectively improve the recognition performance of end-to-end automatic speech recognition (ASR) systems…
cs.SD2023★ 12 cited
LightGrad: Lightweight Diffusion Probabilistic Model for Text-to-Speech
Jie Chen, Xingchen Song, Zhendong Peng +3
Recent advances in neural text-to-speech (TTS) models bring thousands of TTS applications into daily life, where models are deployed in cloud to provide services for customs. Among…