7 citations · 14 across the 5 of their papers we have counts for
4 papers · 1 filter
Federated Learning in ASR: Not as Easy as You Think
Wentao Yu, Jan Freiwald, Sören Tewes +2
With the growing availability of smart devices and cloud services, personal speech assistance systems are increasingly used on a daily basis. Most devices redirect the voice record…
Large-vocabulary Audio-visual Speech Recognition in Noisy Environments
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
Audio-visual speech recognition (AVSR) can effectively and significantly improve the recognition rates of small-vocabulary systems, compared to their audio-only counterparts. For l…
Fusing information streams in end-to-end audio-visual speech recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
End-to-end acoustic speech recognition has quickly gained widespread popularity and shows promising results in many studies. Specifically the joint transformer/CTC model provides v…
Multimodal Integration for Large-Vocabulary Audio-Visual Speech Recognition
Wentao Yu, Steffen Zeiler, Dorothea Kolossa
For many small- and medium-vocabulary tasks, audio-visual speech recognition can significantly improve the recognition rates compared to audio-only systems. However, there is still…