9 papers
Why Do You Say It Like That? A Phoneme-Level Framework for Explainable Speech Deepfake Detection
Anna Taylor, Michele Panariello, Massimiliano Todisco +3
As the accuracy of speech deepfake detection improves with the use of self-supervised representations such as wav2vec 2.0 and HuBERT, understanding why the speech is classified as…
Positive-Incentive Noise Predictor for Adversarial Purification in Speaker Verification
Yibo Bai, Sizhou Chen, Michele Panariello +5
Modern automatic speaker verification (ASV) systems are vulnerable to adversarial perturbations. Diffusion-based purification has recently shown strong effectiveness against such p…
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
Emma Coletta, Massimiliano Todisco, Michele Panariello +2
We introduce Latent Secret Spin (LSS), a blind speech watermarking method based on geometric operations in codec latent space. Based upon orthogonal rotations to principal componen…
Evaluating voice anonymisation using similarity rank disclosure
Shilpa Chandra, Matteo Pettenò, Nicholas Evans +7
The evaluation of voice anonymisation remains challenging. Current practice relies on automatic speaker verification metrics such as the equal error rate (EER). Performance estimat…
Identity leakage through accent cues in voice anonymisation
Rayane Bakari, Olivier Le Blouch, Nicolas Gengembre +2
Voice anonymisation is used to conceal voice identity while preserving linguistic content. Even if anonymisation seems strong, non-timbral cues such as accent that remain post-anon…
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
Natalia Tomashenko, Xiaoxiao Miao, Pierre Champion +7
We present results and analyses from the third VoicePrivacy Challenge held in 2024, which focuses on advancing voice anonymization technologies. The task was to develop a voice ano…