8 papers
Feasibility of Time-Domain DNN-Based Speech Enhancement on Embedded FPGA for Hearing Aids
Feyisayo Olalere, Umut Altin, Kiki van der Heijden +1
Hearing aids impose strict latency and power constraints that current DNN-based speech enhancement systems struggle to meet on embedded hardware. We characterize this gap by deploy…
GRAM: Spatial general-purpose audio representations for real-world environments
Goksenin Yuksel, Marcel van Gerven, Kiki van der Heijden
Audio foundation models learn general-purpose audio representations that facilitate a wide range of downstream tasks. While the performance of these models has greatly increased fo…
GRAM: Spatial general-purpose audio representation models for real-world applications
Goksenin Yuksel, Marcel van Gerven, Kiki van der Heijden
Audio foundation models learn general-purpose audio representations that facilitate a wide range of downstream tasks. While the performance of these models has greatly increased fo…
Speech Separation for Hearing-Impaired Children in the Classroom
Feyisayo Olalere, Kiki van der Heijden, H. Christiaan Stronks +3
Classroom environments are particularly challenging for children with hearing impairments, where background noise, multiple talkers, and reverberation degrade speech perception. Th…
WavJEPA: Semantic learning unlocks robust audio foundation models for raw waveforms
Goksenin Yuksel, Pierre Guetschel, Michael Tangermann +2
Learning audio representations from raw waveforms overcomes key limitations of spectrogram-based audio representation learning, such as the long latency of spectrogram computation…
Leveraging Spatial Cues from Cochlear Implant Microphones to Efficiently Enhance Speech Separation in Real-World Listening Scenes
Feyisayo Olalere, Kiki van der Heijden, Christiaan H. Stronks +3
Speech separation approaches for single-channel, dry speech mixtures have significantly improved. However, real-world spatial and reverberant acoustic environments remain challengi…