16 citations · 44 across the 18 of their papers we have counts for
10 papers · 1 filter
Federated Learning in ASR: Not as Easy as You Think
Wentao Yu, Jan Freiwald, Sören Tewes +2
With the growing availability of smart devices and cloud services, personal speech assistance systems are increasingly used on a daily basis. Most devices redirect the voice record…
Large-vocabulary Audio-visual Speech Recognition in Noisy Environments
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
Audio-visual speech recognition (AVSR) can effectively and significantly improve the recognition rates of small-vocabulary systems, compared to their audio-only counterparts. For l…
O2D2: Out-Of-Distribution Detector to Capture Undecidable Trials in Authorship Verification
Benedikt Boenninghoff, Robert M. Nickel, Dorothea Kolossa
The PAN 2021 authorship verification (AV) challenge is part of a three-year strategy, moving from a cross-topic/closed-set AV task to a cross-topic/open-set AV task over a collecti…
Self-Calibrating Neural-Probabilistic Model for Authorship Verification Under Covariate Shift
Benedikt Boenninghoff, Dorothea Kolossa, Robert M. Nickel
We are addressing two fundamental problems in authorship verification (AV): Topic variability and miscalibration. Variations in the topic of two disputed texts are a major cause of…
PILOT: Introducing Transformers for Probabilistic Sound Event Localization
Christopher Schymura, Benedikt Bönninghoff, Tsubasa Ochiai +5
Sound event localization aims at estimating the positions of sound sources in the environment with respect to an acoustic receiver (e.g. a microphone array). Recent advances in thi…
Fusing information streams in end-to-end audio-visual speech recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
End-to-end acoustic speech recognition has quickly gained widespread popularity and shows promising results in many studies. Specifically the joint transformer/CTC model provides v…