Showing eess.ASShow all
3 papers · 1 filter
eess.AS2025
IDMap: A Pseudo-Speaker Generator Framework Based on Speaker Identity Index to Vector Mapping
Zeyan Liu, Liping Chen, Kong Aik Lee +1
Facilitated by the speech generation framework that disentangles speech into content, speaker, and prosody, voice anonymization is accomplished by substituting the original speaker…
eess.AS2025
Pinhole Effect on Linkability and Dispersion in Speaker Anonymization
Kong Aik Lee, Zeyan Liu, Liping Chen +1
Speaker anonymization aims to conceal speaker-specific attributes in speech signals, making the anonymized speech unlinkable to the original speaker identity. Recent approaches ach…
eess.AS2024
Exploring Audio-Visual Information Fusion for Sound Event Localization and Detection In Low-Resource Realistic Scenarios
Ya Jiang, Qing Wang, Jun Du +9
This study presents an audio-visual information fusion approach to sound event localization and detection (SELD) in low-resource scenarios. We aim at utilizing audio and video moda…