2 papers
cs.SD2019
Analysis of Deep Clustering as Preprocessing for Automatic Speech Recognition of Sparsely Overlapping Speech
Tobias Menne, Ilya Sklyar, Ralf Schlüter +1
Significant performance degradation of automatic speech recognition (ASR) systems is observed when the audio signal contains cross-talk. One of the recently proposed approaches to…
cs.CL2018
Speaker Adapted Beamforming for Multi-Channel Automatic Speech Recognition
Tobias Menne, Ralf Schlüter, Hermann Ney
This paper presents, in the context of multi-channel ASR, a method to adapt a mask based, statistically optimal beamforming approach to a speaker of interest. The beamforming vecto…