2 papers
eess.AS2024
Attention-Based Audio Embeddings for Query-by-Example
Anup Singh, Kris Demuynck, Vipul Arora
An ideal audio retrieval system efficiently and robustly recognizes a short query snippet from an extensive database. However, the performance of well-known audio fingerprinting sy…
eess.AS2024
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
Jenthe Thienpondt, Kris Demuynck
Current speaker diarization systems rely on an external voice activity detection model prior to speaker embedding extraction on the detected speech segments. In this paper, we esta…