16 citations · 33 across the 12 of their papers we have counts for
5 papers · 1 filter
Unsupervised Classification of Voiced Speech and Pitch Tracking Using Forward-Backward Kalman Filtering
Benedikt Boenninghoff, Robert M. Nickel, Steffen Zeiler +1
The detection of voiced speech, the estimation of the fundamental frequency, and the tracking of pitch values over time are crucial subtasks for a variety of speech processing tech…
Exploiting Attention-based Sequence-to-Sequence Architectures for Sound Event Localization
Christopher Schymura, Tsubasa Ochiai, Marc Delcroix +4
Sound event localization frameworks based on deep neural networks have shown increased robustness with respect to reverberation and noise in comparison to classical parametric appr…
Data Fusion for Audiovisual Speaker Localization: Extending Dynamic Stream Weights to the Spatial Domain
Julio Wissing, Benedikt Boenninghoff, Dorothea Kolossa +6
Estimating the positions of multiple speakers can be helpful for tasks like automatic speech recognition or speaker diarization. Both applications benefit from a known speaker posi…
Joining Sound Event Detection and Localization Through Spatial Segregation
Ivo Trowitzsch, Christopher Schymura, Dorothea Kolossa +1
Identification and localization of sounds are both integral parts of computational auditory scene analysis. Although each can be solved separately, the goal of forming coherent aud…
An Active Machine Hearing System for Auditory Stream Segregation
Christopher Schymura, Thomas Walther, Dorothea Kolossa
This study describes a binaural machine hearing system that is capable of performing auditory stream segregation in scenarios where multiple sound sources are present. The process…