29 citations · 42 across the 4 of their papers we have counts for
1 paper · 1 filter
Youssef Mroueh, Etienne Marcheret, Vaibhava Goel
In this paper, we present methods in deep multimodal learning for fusing speech and visual modalities for Audio-Visual Automatic Speech Recognition (AV-ASR). First, we study an app…