31 citations · 77 across the 9 of their papers we have counts for
3 papers · 1 filter
Recurrent Neural Network Transducer for Audio-Visual Speech Recognition
Takaki Makino, Hank Liao, Yannis Assael +4
This work presents a large-scale audio-visual speech recognition system based on a recurrent neural network transducer (RNN-T) architecture. To support the development of such a sy…
Restoring ancient text using deep learning: a case study on Greek epigraphy
Yannis Assael, Thea Sommerschield, Jonathan Prag
Ancient history relies on disciplines such as epigraphy, the study of ancient inscribed texts, for evidence of the recorded past. However, these texts, "inscriptions", are often da…
Speech bandwidth extension with WaveNet
Archit Gupta, Brendan Shillingford, Yannis Assael +1
Large-scale mobile communication systems tend to contain legacy transmission channels with narrowband bottlenecks, resulting in characteristic "telephone-quality" audio. While high…