402 citations · 438 across the 13 of their papers we have counts for
13 papers · 1 filter
Adaptive Slimming for Scalable and Efficient Speech Enhancement
Riccardo Miccini, Minje Kim, Clément Laroche +2
Speech enhancement (SE) enables robust speech recognition, real-time communication, hearing aids, and other applications where speech quality is crucial. However, deploying such sy…
Generative Data Augmentation Challenge: Synthesis of Room Acoustics for Speaker Distance Estimation
Jackie Lin, Georg Götz, Hermes Sampedro Llopis +9
This paper describes the synthesis of the room acoustics challenge as a part of the generative data augmentation workshop at ICASSP 2025. The challenge defines a unique generative…
Rethinking Non-Negative Matrix Factorization with Implicit Neural Representations
Krishna Subramani, Paris Smaragdis, Takuya Higuchi +1
Non-negative Matrix Factorization (NMF) is a powerful technique for analyzing regularly-sampled data, i.e., data that can be stored in a matrix. For audio, this has led to numerous…
Mechatronic Generation of Datasets for Acoustics Research
Austin Lu, Ethaniel Moore, Arya Nallanthighall +5
We address the challenge of making spatial audio datasets by proposing a shared mechanized recording space that can run custom acoustic experiments: a Mechatronic Acoustic Research…
Noise-Robust DSP-Assisted Neural Pitch Estimation with Very Low Complexity
Krishna Subramani, Jean-Marc Valin, Jan Buethe +2
Pitch estimation is an essential step of many speech processing algorithms, including speech coding, synthesis, and enhancement. Recently, pitch estimators based on deep neural net…
Real-Time Packet Loss Concealment With Mixed Generative and Predictive Model
Jean-Marc Valin, Ahmed Mustafa, Christopher Montgomery +4
As deep speech enhancement algorithms have recently demonstrated capabilities greatly surpassing their traditional counterparts for suppressing noise, reverberation and echo, atten…