activity
20182022
most citedEnd-to-End Multi-Channel Speech Separation

80 citations · 180 across the 15 of their papers we have counts for

collaborators
Showing cs.SDShow all

6 papers · 1 filter

cs.SD20226 cited

The DKU-Tencent System for the VoxCeleb Speaker Recognition Challenge 2022

Xiaoyi Qin, Na Li, Yuke Lin +4

This paper is the system description of the DKU-Tencent System for the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC22). In this challenge, we focus on track1 and track3. For…

cs.SD20221 cited

3M: Multi-loss, Multi-path and Multi-level Neural Networks for speech recognition

Zhao You, Shulin Feng, Dan Su +1

Recently, Conformer based CTC/AED model has become a mainstream architecture for ASR. In this paper, based on our prior work, we identify and integrate several approaches to achiev…

cs.SD20213 cited

Simple Attention Module based Speaker Verification with Iterative noisy label detection

Xiaoyi Qin, Na Li, Chao Weng +2

Recently, the attention mechanism such as squeeze-and-excitation module (SE) and convolutional block attention module (CBAM) has achieved great success in deep learning-based speak…

cs.SD20211 cited

Raw Waveform Encoder with Multi-Scale Globally Attentive Locally Recurrent Networks for End-to-End Speech Recognition

Max W. Y. Lam, Jun Wang, Chao Weng +2

End-to-end speech recognition generally uses hand-engineered acoustic features as input and excludes the feature extraction module from its joint optimization. To extract learnable…

cs.SD201980 cited

End-to-End Multi-Channel Speech Separation

Rongzhi Gu, Jian Wu, Shi-Xiong Zhang +6

The end-to-end approach for single-channel speech separation has been studied recently and shown promising results. This paper extended the previous approach and proposed a new end…

cs.SD2018

Deep Extractor Network for Target Speaker Recovery From Single Channel Speech Mixtures

Jun Wang, Jie Chen, Dan Su +4

Speaker-aware source separation methods are promising workarounds for major difficulties such as arbitrary source permutation and unknown number of sources. However, it remains cha…