4 citations · 13 across the 15 of their papers we have counts for
15 papers
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
Michael Kuhlmann, Tobias Cord-Landwehr, Reinhold Haeb-Umbach
Automatic subjective speech quality assessment (SSQA) traditionally estimates speech quality on an utterance or system level. While this resolution was adequate for older transmiss…
Loose coupling of spectral and spatial models for multi-channel diarization and enhancement of meetings in dynamic environments
Adrian Meise, Tobias Cord-Landwehr, Christoph Boeddeker +3
Sound capture by microphone arrays opens the possibility to exploit spatial, in addition to spectral, information for diarization and signal enhancement, two important tasks in mee…
On the Role of Spatial Features in Foundation-Model-Based Speaker Diarization
Marc Deegen, Tobias Gburrek, Tobias Cord-Landwehr +4
Recent advances in speaker diarization exploit large pretrained foundation models, such as WavLM, to achieve state-of-the-art performance on multiple datasets. Systems like DiariZe…
On the Application of Diffusion Models for Simultaneous Denoising and Dereverberation
Adrian Meise, Tobias Cord-Landwehr, Reinhold Haeb-Umbach
Diffusion models have been shown to achieve natural-sounding enhancement of speech degraded by noise or reverberation. However, their simultaneous denoising and dereverberation cap…
Spatio-spectral diarization of meetings by combining TDOA-based segmentation and speaker embedding-based clustering
Tobias Cord-Landwehr, Tobias Gburrek, Marc Deegen +1
We propose a spatio-spectral, combined model-based and data-driven diarization pipeline consisting of TDOA-based segmentation followed by embedding-based clustering. The proposed s…
Simultaneous Diarization and Separation of Meetings through the Integration of Statistical Mixture Models
Tobias Cord-Landwehr, Christoph Boeddeker, Reinhold Haeb-Umbach
We propose an approach for simultaneous diarization and separation of meeting data. It consists of a complex Angular Central Gaussian Mixture Model (cACGMM) for speech source separ…