4 citations · 6 across the 2 of their papers we have counts for
4 papers
Cross-Attention is all you need: Real-Time Streaming Transformers for Personalised Speech Enhancement
Shucong Zhang, Malcolm Chadwick, Alberto Gil C. P. Ramos +1
Personalised speech enhancement (PSE), which extracts only the speech of a target user and removes everything else from a recorded audio clip, can potentially improve users' experi…
The Benefit of the Doubt: Uncertainty Aware Sensing for Edge Computing Platforms
Lorena Qendro, Jagmohan Chauhan, Alberto Gil C. P. Ramos +1
Neural networks (NNs) lack measures of "reliability" estimation that would enable reasoning over their predictions. Despite the vital importance, especially in areas of human well-…
Bunched LPCNet : Vocoder for Low-cost Neural Text-To-Speech Systems
Ravichander Vipperla, Sangjun Park, Kihyun Choo +6
LPCNet is an efficient vocoder that combines linear prediction and deep neural network modules to keep the computational complexity low. In this work, we present two techniques to…
Iterative Compression of End-to-End ASR Model using AutoML
Abhinav Mehrotra, Łukasz Dudziak, Jinsu Yeo +9
Increasing demand for on-device Automatic Speech Recognition (ASR) systems has resulted in renewed interests in developing automatic model compression techniques. Past research hav…