42 citations · 119 across the 12 of their papers we have counts for
8 papers · 1 filter
Multi-view Audio and Music Classification
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
We propose in this work a multi-view learning approach for audio and music classification. Considering four typical low-level representations (i.e. different views) commonly used f…
Inception-Based Network and Multi-Spectrogram Ensemble Applied For Predicting Respiratory Anomalies and Lung Diseases
Lam Pham, Huy Phan, Ross King +2
This paper presents an inception-based deep neural network for detecting lung diseases using respiratory sound input. Recordings of respiratory sound collected from patients are fi…
Self-Attention Generative Adversarial Network for Speech Enhancement
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
Existing generative adversarial networks (GANs) for speech enhancement solely rely on the convolution operation, which may obscure temporal dependencies across the sequence input.…
Robust Acoustic Scene Classification using a Multi-Spectrogram Encoder-Decoder Framework
Lam Pham, Huy Phan, Truc Nguyen +3
This article proposes an encoder-decoder network model for Acoustic Scene Classification (ASC), the task of identifying the scene of an audio recording from its acoustic signature.…
Robust Deep Learning Framework For Predicting Respiratory Anomalies and Diseases
Lam Pham, Ian McLoughlin, Huy Phan +3
This paper presents a robust deep learning framework developed to detect respiratory diseases from recordings of respiratory sounds. The complete detection process firstly involves…
Spatio-Temporal Attention Pooling for Audio Scene Classification
Huy Phan, Oliver Y. Chén, Lam Pham +4
Acoustic scenes are rich and redundant in their content. In this work, we present a spatio-temporal attention pooling layer coupled with a convolutional recurrent neural network to…