activity
20172023
most citedIntegrating the Data Augmentation Scheme with Various Classifiers for Acoustic Scene Modeling

67 citations · 89 across the 14 of their papers we have counts for

collaborators
Showing eess.ASShow all

7 papers · 1 filter

eess.AS20232 cited

Alternative Pseudo-Labeling for Semi-Supervised Automatic Speech Recognition

Han Zhu, Dongji Gao, Gaofeng Cheng +3

When labeled data is insufficient, semi-supervised learning with the pseudo-labeling technique can significantly improve the performance of automatic speech recognition. However, p…

eess.AS2023

Online Hybrid CTC/Attention End-to-End Automatic Speech Recognition Architecture

Haoran Miao, Gaofeng Cheng, Pengyuan Zhang +1

Recently, there has been increasing progress in end-to-end automatic speech recognition (ASR) architecture, which transcribes speech to text without any pre-trained alignments. One…

eess.AS20218 cited

Improved Conformer-based End-to-End Speech Recognition Using Neural Architecture Search

Yukun Liu, Ta Li, Pengyuan Zhang +1

Recently neural architecture search(NAS) has been successfully used in image classification, natural language processing, and automatic speech recognition(ASR) tasks for finding th…

eess.AS20201 cited

Multi-Accent Adaptation based on Gate Mechanism

Han Zhu, Li Wang, Pengyuan Zhang +1

When only a limited amount of accented speech data is available, to promote multi-accent speech recognition performance, the conventional approach is accent-specific adaptation, wh…

eess.AS20203 cited

Transformer-based Online CTC/attention End-to-End Speech Recognition Architecture

Haoran Miao, Gaofeng Cheng, Changfeng Gao +2

Recently, Transformer has gained success in automatic speech recognition (ASR) field. However, it is challenging to deploy a Transformer-based end-to-end (E2E) model for online spe…

eess.AS20191 cited

Multi-Talker MVDR Beamforming Based on Extended Complex Gaussian Mixture Model

Hangting Chen, Pengyuan Zhang, Yonghong Yan

In this letter, we present a novel multi-talker minimum variance distortionless response (MVDR) beamforming as the front-end of an automatic speech recognition (ASR) system in a di…