42 citations · 110 across the 9 of their papers we have counts for
19 papers
Polyphonic audio event detection: multi-label or multi-class multi-task classification problem?
Huy Phan, Thi Ngoc Tho Nguyen, Philipp Koch +1
Polyphonic events are the main error source of audio event detection (AED) systems. In deep-learning context, the most common approach to deal with event overlaps is to treat the A…
Multi-view Audio and Music Classification
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
We propose in this work a multi-view learning approach for audio and music classification. Considering four typical low-level representations (i.e. different views) commonly used f…
Inception-Based Network and Multi-Spectrogram Ensemble Applied For Predicting Respiratory Anomalies and Lung Diseases
Lam Pham, Huy Phan, Ross King +2
This paper presents an inception-based deep neural network for detecting lung diseases using respiratory sound input. Recordings of respiratory sound collected from patients are fi…
Self-Attention Generative Adversarial Network for Speech Enhancement
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
Existing generative adversarial networks (GANs) for speech enhancement solely rely on the convolution operation, which may obscure temporal dependencies across the sequence input.…
On Multitask Loss Function for Audio Event Detection and Localization
Huy Phan, Lam Pham, Philipp Koch +3
Audio event localization and detection (SELD) have been commonly tackled using multitask models. Such a model usually consists of a multi-label event classification branch with sig…
XSleepNet: Multi-View Sequential Model for Automatic Sleep Staging
Huy Phan, Oliver Y. Chén, Minh C. Tran +3
Automating sleep staging is vital to scale up sleep assessment and diagnosis to serve millions experiencing sleep deprivation and disorders and enable longitudinal sleep monitoring…