31 citations · 62 across the 6 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2019★ 15 cited
Recurrent Neural Network Transducer for Audio-Visual Speech Recognition
Takaki Makino, Hank Liao, Yannis Assael +4
This work presents a large-scale audio-visual speech recognition system based on a recurrent neural network transducer (RNN-T) architecture. To support the development of such a sy…
eess.AS2019
Speech bandwidth extension with WaveNet
Archit Gupta, Brendan Shillingford, Yannis Assael +1
Large-scale mobile communication systems tend to contain legacy transmission channels with narrowband bottlenecks, resulting in characteristic "telephone-quality" audio. While high…