25 citations · 50 across the 11 of their papers we have counts for
9 papers · 1 filter
Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
Fan Yu, Shiliang Zhang, Pengcheng Guo +13
The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Grand Challenge (M2MeT) focuses on one of the most valuable and the most challenging scenarios of speech technologie…
Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction
Heming Wang, Yao Qian, Xiaofei Wang +6
Noise robustness is essential for deploying automatic speech recognition (ASR) systems in real-world environments. One way to reduce the effect of noise interference is to employ a…
VoiceFixer: Toward General Speech Restoration with Neural Vocoder
Haohe Liu, Qiuqiang Kong, Qiao Tian +4
Speech restoration aims to remove distortions in speech signals. Prior methods mainly focus on single-task speech restoration (SSR), such as speech denoising or speech declipping.…
Localization Based Sequential Grouping for Continuous Speech Separation
Zhong-Qiu Wang, DeLiang Wang
This study investigates robust speaker localization for con-tinuous speech separation and speaker diarization, where we use speaker directions to group non-contiguous segments of t…
Speaker Separation Using Speaker Inventories and Estimated Speech
Peidong Wang, Zhuo Chen, DeLiang Wang +2
We propose speaker separation using speaker inventories and estimated speech (SSUSIES), a framework leveraging speaker profiles and estimated speech for speaker separation. SSUSIES…
On Cross-Corpus Generalization of Deep Learning Based Speech Enhancement
Ashutosh Pandey, DeLiang Wang
In recent years, supervised approaches using deep neural networks (DNNs) have become the mainstream for speech enhancement. It has been established that DNNs generalize well to unt…