16 citations · 44 across the 8 of their papers we have counts for
17 papers
An Ensemble of Deep Learning Frameworks Applied For Predicting Respiratory Anomalies
Lam Pham, Dat Ngo, Truong Hoang +2
In this paper, we evaluate various deep learning frameworks for detecting respiratory anomalies from input audio recordings. To this end, we firstly transform audio respiratory cyc…
Extremely Low Footprint End-to-End ASR System for Smart Device
Zhifu Gao, Yiwu Yao, Shiliang Zhang +3
Recently, end-to-end (E2E) speech recognition has become popular, since it can integrate the acoustic, pronunciation and language models into a single neural network, which outperf…
Multi-view Audio and Music Classification
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
We propose in this work a multi-view learning approach for audio and music classification. Considering four typical low-level representations (i.e. different views) commonly used f…
Universal ASR: Unifying Streaming and Non-Streaming ASR Using a Single Encoder-Decoder Model
Zhifu Gao, Shiliang Zhang, Ming Lei +1
Recently, online end-to-end ASR has gained increasing attention. However, the performance of online systems still lags far behind that of offline systems, with a large gap in quali…
Incandescent Bulb and LED Brake Lights:Novel Analysis of Reaction Times
Ramaswamy Palaniappan, Surej Mouli, Evangelina Fringi +2
Rear-end collision accounts for around 8% of all vehicle crashes in the UK, with the failure to notice or react to a brake light signal being a major contributory cause. Meanwhile…
Self-Attention Generative Adversarial Network for Speech Enhancement
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
Existing generative adversarial networks (GANs) for speech enhancement solely rely on the convolution operation, which may obscure temporal dependencies across the sequence input.…