35 citations · 35 across the 2 of their papers we have counts for
4 papers
Modular End-to-end Automatic Speech Recognition Framework for Acoustic-to-word Model
Qi Liu, Zhehuai Chen, Hao Li +3
End-to-end (E2E) systems have played a more and more important role in automatic speech recognition (ASR) and achieved great performance. However, E2E systems recognize output word…
Sequence Discriminative Training for Deep Learning based Acoustic Keyword Spotting
Zhehuai Chen, Yanmin Qian, Kai Yu
Speech recognition is a sequence prediction problem. Besides employing various deep learning approaches for framelevel classification, sequence-level discriminative training has be…
On Modular Training of Neural Acoustics-to-Word Model for LVCSR
Zhehuai Chen, Qi Liu, Hao Li +1
End-to-end (E2E) automatic speech recognition (ASR) systems directly map acoustics to words using a unified model. Previous works mostly focus on E2E training a single model which…
Progressive Joint Modeling in Unsupervised Single-channel Overlapped Speech Recognition
Zhehuai Chen, Jasha Droppo, Jinyu Li +1
Unsupervised single-channel overlapped speech recognition is one of the hardest problems in automatic speech recognition (ASR). Permutation invariant training (PIT) is a state of t…