1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Richard Rose, Olivier Siohan
This paper presents a new approach for end-to-end audio-visual multi-talker speech recognition. The approach, referred to here as the visual context attention model (VCAM), is impo…