7 citations · 21 across the 13 of their papers we have counts for
17 papers
Cloud-Boosted Low-Compute Multi-Channel Speech Enhancement
Xulin Fan, Juan Azcarreta, Ashutosh Pandey +5
Low-latency, low-compute speech enhancement is essential for wearable devices with real-time communication requirements, but strict computational constraints significantly limit on…
Masked Autoencoders as Universal Speech Enhancer
Rajalaxmi Rajagopalan, Ritwik Giri, Zhiqiang Tang +1
Supervised speech enhancement methods have been very successful. However, in practical scenarios, there is a lack of clean speech, and self-supervised learning-based (SSL) speech e…
Real-time Stereo Speech Enhancement with Spatial-Cue Preservation based on Dual-Path Structure
Masahito Togami, Jean-Marc Valin, Karim Helwani +3
We introduce a real-time, multichannel speech enhancement algorithm which maintains the spatial cues of stereo recordings including two speech sources. Recognizing that each source…
A Framework for Unified Real-time Personalized and Non-Personalized Speech Enhancement
Zhepei Wang, Ritwik Giri, Devansh Shah +3
In this study, we present an approach to train a single speech enhancement network that can perform both personalized and non-personalized speech enhancement. This is achieved by i…
Semi-supervised Time Domain Target Speaker Extraction with Attention
Zhepei Wang, Ritwik Giri, Shrikant Venkataramani +5
In this work, we propose Exformer, a time-domain architecture for target speaker extraction. It consists of a pre-trained speaker embedder network and a separator network based on…
To Dereverb Or Not to Dereverb? Perceptual Studies On Real-Time Dereverberation Targets
Jean-Marc Valin, Ritwik Giri, Shrikant Venkataramani +2
In real life, room effect, also known as room reverberation, and the present background noise degrade the quality of speech. Recently, deep learning-based speech enhancement approa…