5 citations · 7 across the 2 of their papers we have counts for
2 papers
eess.AS2021★ 2 cited
Fusing information streams in end-to-end audio-visual speech recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
End-to-end acoustic speech recognition has quickly gained widespread popularity and shows promising results in many studies. Specifically the joint transformer/CTC model provides v…
eess.AS2020★ 5 cited
Multimodal Integration for Large-Vocabulary Audio-Visual Speech Recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
For many small- and medium-vocabulary tasks, audio-visual speech recognition can significantly improve the recognition rates compared to audio-only systems. However, there is still…