29 citations · 38 across the 6 of their papers we have counts for
1 paper · 2 filters
Youssef Mroueh, Etienne Marcheret, Vaibhava Goel
In this paper, we present methods in deep multimodal learning for fusing speech and visual modalities for Audio-Visual Automatic Speech Recognition (AV-ASR). First, we study an app…