18 citations · 41 across the 29 of their papers we have counts for
24 papers
Explaining Speaker and Spoof Embeddings via Probing
Xuechen Liu, Junichi Yamagishi, Md Sahidullah +1
This study investigates the explainability of embedding representations, specifically those used in modern audio spoofing detection systems based on deep neural networks, known as…
Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data
Yun Liu, Xuechen Liu, Xiaoxiao Miao +1
Target speaker extraction (TSE) is essential in speech processing applications, particularly in scenarios with complex acoustic environments. Current TSE systems face challenges in…
It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model
Mingyi Shi, Dafei Qin, Leo Ho +4
Conversational scenarios are very common in real-world settings, yet existing co-speech motion synthesis approaches often fall short in these contexts, where one person's audio and…
A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection
Xuechen Liu, Xin Wang, Junichi Yamagishi
Audio spoofing detection has become increasingly important due to the rise in real-world cases. Current spoofing detectors, referred to as spoofing countermeasures (CM), are mainly…
ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
Xin Wang, Hector Delgado, Hemlata Tak +10
ASVspoof 5 is the fifth edition in a series of challenges that promote the study of speech spoofing and deepfake attacks, and the design of detection solutions. Compared to previou…
The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
Michele Panariello, Natalia Tomashenko, Xin Wang +7
The VoicePrivacy Challenge promotes the development of voice anonymisation solutions for speech technology. In this paper we present a systematic overview and analysis of the secon…