19 citations
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
Noise-Robust Target-Speaker Voice Activity Detection Through Self-Supervised Pretraining
Holger Severin Bovbjerg, Jan Østergaard, Jesper Jensen +1
Target-Speaker Voice Activity Detection (TS-VAD) is the task of detecting the presence of speech from a known target-speaker in an audio frame. Recently, deep neural network-based…
eess.AS2024★ 4 cited
How to train your ears: Auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing losses
Peter Leer, Jesper Jensen, Zheng-Hua Tan +2
Advanced auditory models are useful in designing signal-processing algorithms for hearing-loss compensation or speech enhancement. Such auditory models provide rich and detailed de…