activity
20202022
most citedChannel-wise Gated Res2Net: Towards Robust Detection of Synthetic Speech Attacks

6 citations · 6 across the 4 of their papers we have counts for

collaborators

5 papers

eess.AS2022

Speaker Identity Preservation in Dysarthric Speech Reconstruction by Adversarial Speaker Adaptation

Disong Wang, Songxiang Liu, Xixin Wu +4

Dysarthric speech reconstruction (DSR), which aims to improve the quality of dysarthric speech, remains a challenge, not only because we need to restore the speech to be normal, bu…

eess.AS2022

The CUHK-TENCENT speaker diarization system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge

Naijun Zheng, Na Li, Xixin Wu +6

This paper describes our speaker diarization system submitted to the Multi-channel Multi-party Meeting Transcription (M2MeT) challenge, where Mandarin meeting data were recorded in…

eess.AS20216 cited

Channel-wise Gated Res2Net: Towards Robust Detection of Synthetic Speech Attacks

Xu Li, Xixin Wu, Hui Lu +2

Existing approaches for anti-spoofing in automatic speaker verification (ASV) still lack generalizability to unseen attacks. The Res2Net approach designs a residual-like connection…

cs.SD2021

VAENAR-TTS: Variational Auto-Encoder based Non-AutoRegressive Text-to-Speech Synthesis

Hui Lu, Zhiyong Wu, Xixin Wu +4

This paper describes a variational auto-encoder based non-autoregressive text-to-speech (VAENAR-TTS) model. The autoregressive TTS (AR-TTS) models based on the sequence-to-sequence…

eess.AS2020

Learning Explicit Prosody Models and Deep Speaker Embeddings for Atypical Voice Conversion

Disong Wang, Songxiang Liu, Lifa Sun +3

Though significant progress has been made for the voice conversion (VC) of typical speech, VC for atypical speech, e.g., dysarthric and second-language (L2) speech, remains a chall…