collaborators

6 papers

cs.SD2026

DeepFense: A Unified, Modular, and Extensible Framework for Robust Deepfake Audio Detection

Yassine El Kheir, Arnab Das, Yixuan Xiao +6

Speech deepfake detection is a well-established research field with different models, datasets, and training strategies. However, the lack of standardized implementations and evalu…

eess.AS2025

Use Cases for Voice Anonymization

Sarina Meyer, Ngoc Thang Vu

The performance of a voice anonymization system is typically measured according to its ability to hide the speaker's identity and keep the data's utility for downstream tasks. This…

eess.AS2025

The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation

Michele Panariello, Sarina Meyer, Pierre Champion +4

Voice anonymisation aims to conceal the voice identity of speakers in speech recordings. Privacy protection is usually estimated from the difficulty of using a speaker verification…

eess.AS2025

First Steps Towards Voice Anonymization for Code-Switching Speech

Sarina Meyer, Ekaterina Kolos, Ngoc Thang Vu

The goal of voice anonymization is to modify an audio such that the true identity of its speaker is hidden. Research on this task is typically limited to the same English read spee…

eess.AS2025

Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis

Paul Mayer, Florian Lux, Alejandro Pérez-González-de-Martos +4

While generative methods have progressed rapidly in recent years, generating expressive prosody for an utterance remains a challenging task in text-to-speech synthesis. This is par…

cs.SD2025

High-Resolution Speech Restoration with Latent Diffusion Model

Tushar Dhyani, Florian Lux, Michele Mancusi +3

Traditional speech enhancement methods often oversimplify the task of restoration by focusing on a single type of distortion. Generative models that handle multiple distortions fre…