26 citations · 67 across the 14 of their papers we have counts for
5 papers · 1 filter
On Word Error Rate Definitions and their Efficient Computation for Multi-Speaker Speech Recognition Systems
Thilo von Neumann, Christoph Boeddeker, Keisuke Kinoshita +2
We propose a general framework to compute the word error rate (WER) of ASR systems that process recordings containing multiple speakers at their input and that produce multiple out…
MMS-MSG: A Multi-purpose Multi-Speaker Mixture Signal Generator
Tobias Cord-Landwehr, Thilo von Neumann, Christoph Boeddeker +1
The scope of speech enhancement has changed from a monolithic view of single, independent tasks, to a joint processing of complex conversational speech recordings. Training and eva…
Utterance-by-utterance overlap-aware neural diarization with Graph-PIT
Keisuke Kinoshita, Thilo von Neumann, Marc Delcroix +2
Recent speaker diarization studies showed that integration of end-to-end neural diarization (EEND) and clustering-based diarization is a promising approach for achieving state-of-t…
A Meeting Transcription System for an Ad-Hoc Acoustic Sensor Network
Tobias Gburrek, Christoph Boeddeker, Thilo von Neumann +3
We propose a system that transcribes the conversation of a typical meeting scenario that is captured by a set of initially unsynchronized microphone arrays at unknown positions. It…
An Initialization Scheme for Meeting Separation with Spatial Mixture Models
Christoph Boeddeker, Tobias Cord-Landwehr, Thilo von Neumann +1
Spatial mixture model (SMM) supported acoustic beamforming has been extensively used for the separation of simultaneously active speakers. However, it has hardly been considered fo…