42 citations · 58 across the 4 of their papers we have counts for
12 papers
Polyphonic audio event detection: multi-label or multi-class multi-task classification problem?
Huy Phan, Thi Ngoc Tho Nguyen, Philipp Koch +1
Polyphonic events are the main error source of audio event detection (AED) systems. In deep-learning context, the most common approach to deal with event overlaps is to treat the A…
Multi-view Audio and Music Classification
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
We propose in this work a multi-view learning approach for audio and music classification. Considering four typical low-level representations (i.e. different views) commonly used f…
Self-Attention Generative Adversarial Network for Speech Enhancement
Huy Phan, Huy Le Nguyen, Oliver Y. Chén +4
Existing generative adversarial networks (GANs) for speech enhancement solely rely on the convolution operation, which may obscure temporal dependencies across the sequence input.…
On Multitask Loss Function for Audio Event Detection and Localization
Huy Phan, Lam Pham, Philipp Koch +3
Audio event localization and detection (SELD) have been commonly tackled using multitask models. Such a model usually consists of a multi-label event classification branch with sig…
XSleepNet: Multi-View Sequential Model for Automatic Sleep Staging
Huy Phan, Oliver Y. Chén, Minh C. Tran +3
Automating sleep staging is vital to scale up sleep assessment and diagnosis to serve millions experiencing sleep deprivation and disorders and enable longitudinal sleep monitoring…
Personalized Automatic Sleep Staging with Single-Night Data: a Pilot Study with KL-Divergence Regularization
Huy Phan, Kaare Mikkelsen, Oliver Y. Chén +4
Brain waves vary between people. An obvious way to improve automatic sleep staging for longitudinal sleep monitoring is personalization of algorithms based on individual characteri…