50 citations · 61 across the 3 of their papers we have counts for
4 papers
A Treatise On FST Lattice Based MMI Training
Adnan Haider, Tim Ng, Zhen Huang +2
Maximum mutual information (MMI) has become one of the two de facto methods for sequence-level training of speech recognition acoustic models. This paper aims to isolate, identify…
Data Augmentation For Children's Speech Recognition -- The "Ethiopian" System For The SLT 2021 Children Speech Recognition Challenge
Guoguo Chen, Xingyu Na, Yongqing Wang +4
This paper presents the "Ethiopian" system for the SLT 2021 Children Speech Recognition Challenge. Various data processing and augmentation techniques are proposed to tackle childr…
AISHELL-2: Transforming Mandarin ASR Research Into Industrial Scale
Jiayu Du, Xingyu Na, Xuechen Liu +1
AISHELL-1 is by far the largest open-source speech corpus available for Mandarin speech recognition research. It was released with a baseline system containing solid training and t…
AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline
Hui Bu, Jiayu Du, Xingyu Na +2
An open-source Mandarin speech corpus called AISHELL-1 is released. It is by far the largest corpus which is suitable for conducting the speech recognition research and building sp…