3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CL2022
Open Source MagicData-RAMC: A Rich Annotated Mandarin Conversational(RAMC) Speech Dataset
Zehui Yang, Yifan Chen, Lei Luo +9
This paper introduces a high-quality rich annotated Mandarin conversational (RAMC) speech dataset called MagicData-RAMC. The MagicData-RAMC corpus contains 180 hours of conversatio…
cs.CL2022
Improving CTC-based speech recognition via knowledge transferring from pre-trained language models
Keqi Deng, Songjun Cao, Yike Zhang +4
Recently, end-to-end automatic speech recognition models based on connectionist temporal classification (CTC) have achieved impressive results, especially when fine-tuned from wav2…
cs.CL2017★ 3 cited
An Improved Residual LSTM Architecture for Acoustic Modeling
Lu Huang, Jiasong Sun, Ji Xu +1
Long Short-Term Memory (LSTM) is the primary recurrent neural networks architecture for acoustic modeling in automatic speech recognition systems. Residual learning is an efficient…