5 citations · 9 across the 6 of their papers we have counts for
3 papers · 1 filter
Large-vocabulary Audio-visual Speech Recognition in Noisy Environments
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
Audio-visual speech recognition (AVSR) can effectively and significantly improve the recognition rates of small-vocabulary systems, compared to their audio-only counterparts. For l…
Fusing information streams in end-to-end audio-visual speech recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
End-to-end acoustic speech recognition has quickly gained widespread popularity and shows promising results in many studies. Specifically the joint transformer/CTC model provides v…
Multimodal Integration for Large-Vocabulary Audio-Visual Speech Recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
For many small- and medium-vocabulary tasks, audio-visual speech recognition can significantly improve the recognition rates compared to audio-only systems. However, there is still…