activity
20212023
most citedMulti-Variant Consistency based Self-supervised Learning for Robust Automatic Speech Recognition

3 citations · 5 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL20231 cited

Speech Corpora Divergence Based Unsupervised Data Selection for ASR

Changfeng Gao, Gaofeng Cheng, Pengyuan Zhang +1

Selecting application scenarios matching data is important for the automatic speech recognition (ASR) training, but it is difficult to measure the matching degree of the training c…

cs.CL2022

The Conversational Short-phrase Speaker Diarization (CSSD) Task: Dataset, Evaluation Metric and Baselines

Gaofeng Cheng, Yifan Chen, Runyan Yang +9

The conversation scenario is one of the most important and most challenging scenarios for speech processing technologies because people in conversation respond to each other in a c…

eess.AS20221 cited

Improving Streaming End-to-End ASR on Transformer-based Causal Models with Encoder States Revision Strategies

Zehan Li, Haoran Miao, Keqi Deng +4

There is often a trade-off between performance and latency in streaming automatic speech recognition (ASR). Traditional methods such as look-ahead and chunk-based methods, usually…

eess.AS2022

Interrelate Training and Searching: A Unified Online Clustering Framework for Speaker Diarization

Yifan Chen, Yifan Guo, Qingxuan Li +3

For online speaker diarization, samples arrive incrementally, and the overall distribution of the samples is invisible. Moreover, in most existing clustering-based methods, the tra…

cs.SD20213 cited

Multi-Variant Consistency based Self-supervised Learning for Robust Automatic Speech Recognition

Changfeng Gao, Gaofeng Cheng, Pengyuan Zhang

Automatic speech recognition (ASR) has shown rapid advances in recent years but still degrades significantly in far-field and noisy environments. The recent development of self-sup…