activity
20212024
most citedSpeech Separation Using an Asynchronous Fully Recurrent Convolutional Neural Network

22 citations · 36 across the 17 of their papers we have counts for

collaborators
Showing eess.ASShow all

15 papers · 1 filter

eess.AS2024

Robustness of Speech Separation Models for Similar-pitch Speakers

Bunlong Lay, Sebastian Zaczek, Kristina Tesch +1

Single-channel speech separation is a crucial task for enhancing speech recognition systems in multi-speaker environments. This paper investigates the robustness of state-of-the-ar…

eess.AS2024

EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation

Julius Richter, Yi-Chiao Wu, Steven Krenn +5

We release the EARS (Expressive Anechoic Recordings of Speech) dataset, a high-quality speech dataset comprising 107 speakers from diverse backgrounds, totaling in 100 hours of cle…

eess.AS2024

The PESQetarian: On the Relevance of Goodhart's Law for Speech Enhancement

Danilo de Oliveira, Simon Welker, Julius Richter +1

To obtain improved speech enhancement models, researchers often focus on increasing performance according to specific instrumental metrics. However, when the same metric is used in…

eess.AS2023

Distilling HuBERT with LSTMs via Decoupled Knowledge Distillation

Danilo de Oliveira, Timo Gerkmann

Much research effort is being applied to the task of compressing the knowledge of self-supervised models, which are powerful, yet large and memory consuming. In this work, we show…

eess.AS2023

A Flexible Online Framework for Projection-Based STFT Phase Retrieval

Tal Peer, Simon Welker, Johannes Kolhoff +1

Several recent contributions in the field of iterative STFT phase retrieval have demonstrated that the performance of the classical Griffin-Lim method can be considerably improved…

eess.AS2023

On the Behavior of Intrusive and Non-intrusive Speech Enhancement Metrics in Predictive and Generative Settings

Danilo de Oliveira, Julius Richter, Jean-Marie Lemercier +2

Since its inception, the field of deep speech enhancement has been dominated by predictive (discriminative) approaches, such as spectral mapping or masking. Recently, however, nove…