1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.SD2022★ 1 cited
End-to-end multi-talker audio-visual ASR using an active speaker attention module
Richard Rose, Olivier Siohan
This paper presents a new approach for end-to-end audio-visual multi-talker speech recognition. The approach, referred to here as the visual context attention model (VCAM), is impo…
stat.ML2016
Graph based manifold regularized deep neural networks for automatic speech recognition
Vikrant Singh Tomar, Richard C. Rose
Deep neural networks (DNNs) have been successfully applied to a wide variety of acoustic modeling tasks in recent years. These include the applications of DNNs either in a discrimi…