42 citations · 110 across the 9 of their papers we have counts for
7 papers · 1 filter
Multi-view Audio and Music Classification
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
We propose in this work a multi-view learning approach for audio and music classification. Considering four typical low-level representations (i.e. different views) commonly used f…
Inception-Based Network and Multi-Spectrogram Ensemble Applied For Predicting Respiratory Anomalies and Lung Diseases
Lam Pham, Huy Phan, Ross King +2
This paper presents an inception-based deep neural network for detecting lung diseases using respiratory sound input. Recordings of respiratory sound collected from patients are fi…
Self-Attention Generative Adversarial Network for Speech Enhancement
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
Existing generative adversarial networks (GANs) for speech enhancement solely rely on the convolution operation, which may obscure temporal dependencies across the sequence input.…
Robust Acoustic Scene Classification using a Multi-Spectrogram Encoder-Decoder Framework
Lam Pham, Huy Phan, Truc Nguyen +3
This article proposes an encoder-decoder network model for Acoustic Scene Classification (ASC), the task of identifying the scene of an audio recording from its acoustic signature.…
Spatio-Temporal Attention Pooling for Audio Scene Classification
Huy Phan, Oliver Y. Chén, Lam Pham +4
Acoustic scenes are rich and redundant in their content. In this work, we present a spatio-temporal attention pooling layer coupled with a convolutional recurrent neural network to…
Beyond Equal-Length Snippets: How Long is Sufficient to Recognize an Audio Scene?
Huy Phan, Oliver Y. Chén, Philipp Koch +4
Due to the variability in characteristics of audio scenes, some scenes can naturally be recognized earlier than others. In this work, rather than using equal-length snippets for al…