Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
SEF-MK: Speaker-Embedding-Free Voice Anonymization through Multi-k-means Quantization
Beilong Tang, Xiaoxiao Miao, Xin Wang +1
Voice anonymization protects speaker privacy by concealing identity while preserving linguistic and paralinguistic content. Self-supervised learning (SSL) representations encode li…
cs.SD2024
TSELM: Target Speaker Extraction using Discrete Tokens and Language Models
Beilong Tang, Bang Zeng, Ming Li
We propose TSELM, a novel target speaker extraction network that leverages discrete tokens and language models. TSELM utilizes multiple discretized layers from WavLM as input token…