80 citations · 180 across the 15 of their papers we have counts for
6 papers · 1 filter
The DKU-Tencent System for the VoxCeleb Speaker Recognition Challenge 2022
Xiaoyi Qin, Na Li, Yuke Lin +4
This paper is the system description of the DKU-Tencent System for the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC22). In this challenge, we focus on track1 and track3. For…
3M: Multi-loss, Multi-path and Multi-level Neural Networks for speech recognition
Zhao You, Shulin Feng, Dan Su +1
Recently, Conformer based CTC/AED model has become a mainstream architecture for ASR. In this paper, based on our prior work, we identify and integrate several approaches to achiev…
Simple Attention Module based Speaker Verification with Iterative noisy label detection
Xiaoyi Qin, Na Li, Chao Weng +2
Recently, the attention mechanism such as squeeze-and-excitation module (SE) and convolutional block attention module (CBAM) has achieved great success in deep learning-based speak…
Raw Waveform Encoder with Multi-Scale Globally Attentive Locally Recurrent Networks for End-to-End Speech Recognition
Max W. Y. Lam, Jun Wang, Chao Weng +2
End-to-end speech recognition generally uses hand-engineered acoustic features as input and excludes the feature extraction module from its joint optimization. To extract learnable…
End-to-End Multi-Channel Speech Separation
Rongzhi Gu, Jian Wu, Shi-Xiong Zhang +6
The end-to-end approach for single-channel speech separation has been studied recently and shown promising results. This paper extended the previous approach and proposed a new end…
Deep Extractor Network for Target Speaker Recovery From Single Channel Speech Mixtures
Jun Wang, Jie Chen, Dan Su +4
Speaker-aware source separation methods are promising workarounds for major difficulties such as arbitrary source permutation and unknown number of sources. However, it remains cha…