5 citations · 5 across the 2 of their papers we have counts for
3 papers
eess.AS2021★ 5 cited
Exploring Retraining-Free Speech Recognition for Intra-sentential Code-Switching
Zhen Huang, Xiaodan Zhuang, Daben Liu +3
In this paper, we present our initial efforts for building a code-switching (CS) speech recognition system leveraging existing acoustic models (AMs) and language models (LMs), i.e.…
cs.CL2020
Frame-level SpecAugment for Deep Convolutional Neural Networks in Hybrid ASR Systems
Xinwei Li, Yuanyuan Zhang, Xiaodan Zhuang +1
Inspired by SpecAugment -- a data augmentation method for end-to-end ASR systems, we propose a frame-level SpecAugment method (f-SpecAugment) to improve the performance of deep con…
cs.LG2019
SNDCNN: Self-normalizing deep CNNs with scaled exponential linear units for speech recognition
Zhen Huang, Tim Ng, Leo Liu +3
Very deep CNNs achieve state-of-the-art results in both computer vision and speech recognition, but are difficult to train. The most popular way to train very deep CNNs is to use s…