190 citations · 360 across the 11 of their papers we have counts for
Showing 2019 · eess.ASShow all
2 papers · 2 filters
eess.AS2019★ 15 cited
Recurrent Neural Network Transducer for Audio-Visual Speech Recognition
Takaki Makino, Hank Liao, Yannis Assael +4
This work presents a large-scale audio-visual speech recognition system based on a recurrent neural network transducer (RNN-T) architecture. To support the development of such a sy…
eess.AS2019
Speech bandwidth extension with WaveNet
Archit Gupta, Brendan Shillingford, Yannis Assael +1
Large-scale mobile communication systems tend to contain legacy transmission channels with narrowband bottlenecks, resulting in characteristic "telephone-quality" audio. While high…