67 citations · 89 across the 14 of their papers we have counts for
7 papers · 1 filter
Alternative Pseudo-Labeling for Semi-Supervised Automatic Speech Recognition
Han Zhu, Dongji Gao, Gaofeng Cheng +3
When labeled data is insufficient, semi-supervised learning with the pseudo-labeling technique can significantly improve the performance of automatic speech recognition. However, p…
Online Hybrid CTC/Attention End-to-End Automatic Speech Recognition Architecture
Haoran Miao, Gaofeng Cheng, Pengyuan Zhang +1
Recently, there has been increasing progress in end-to-end automatic speech recognition (ASR) architecture, which transcribes speech to text without any pre-trained alignments. One…
Improved Conformer-based End-to-End Speech Recognition Using Neural Architecture Search
Yukun Liu, Ta Li, Pengyuan Zhang +1
Recently neural architecture search(NAS) has been successfully used in image classification, natural language processing, and automatic speech recognition(ASR) tasks for finding th…
Multi-Accent Adaptation based on Gate Mechanism
Han Zhu, Li Wang, Pengyuan Zhang +1
When only a limited amount of accented speech data is available, to promote multi-accent speech recognition performance, the conventional approach is accent-specific adaptation, wh…
Transformer-based Online CTC/attention End-to-End Speech Recognition Architecture
Haoran Miao, Gaofeng Cheng, Changfeng Gao +2
Recently, Transformer has gained success in automatic speech recognition (ASR) field. However, it is challenging to deploy a Transformer-based end-to-end (E2E) model for online spe…
Multi-Talker MVDR Beamforming Based on Extended Complex Gaussian Mixture Model
Hangting Chen, Pengyuan Zhang, Yonghong Yan
In this letter, we present a novel multi-talker minimum variance distortionless response (MVDR) beamforming as the front-end of an automatic speech recognition (ASR) system in a di…