5 citations · 9 across the 5 of their papers we have counts for
8 papers
Large-vocabulary Audio-visual Speech Recognition in Noisy Environments
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
Audio-visual speech recognition (AVSR) can effectively and significantly improve the recognition rates of small-vocabulary systems, compared to their audio-only counterparts. For l…
Fusing information streams in end-to-end audio-visual speech recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
End-to-end acoustic speech recognition has quickly gained widespread popularity and shows promising results in many studies. Specifically the joint transformer/CTC model provides v…
Unsupervised Classification of Voiced Speech and Pitch Tracking Using Forward-Backward Kalman Filtering
Benedikt Boenninghoff, Robert M. Nickel, Steffen Zeiler +1
The detection of voiced speech, the estimation of the fundamental frequency, and the tracking of pitch values over time are crucial subtasks for a variety of speech processing tech…
Multimodal Integration for Large-Vocabulary Audio-Visual Speech Recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
For many small- and medium-vocabulary tasks, audio-visual speech recognition can significantly improve the recognition rates compared to audio-only systems. However, there is still…
Variational Autoencoder with Embedded Student- Mixture Model for Authorship Attribution
Benedikt Boenninghoff, Steffen Zeiler, Robert M. Nickel +1
Traditional computational authorship attribution describes a classification task in a closed-set scenario. Given a finite set of candidate authors and corresponding labeled texts,…
Similarity Learning for Authorship Verification in Social Media
Benedikt Boenninghoff, Robert M. Nickel, Steffen Zeiler +1
Authorship verification tries to answer the question if two documents with unknown authors were written by the same author or not. A range of successful technical approaches has be…