activity
20182021
most citedCross-task learning for audio tagging, sound event detection and spatial localization: DCASE 2019 baseline systems

36 citations · 69 across the 6 of their papers we have counts for

collaborators

13 papers

eess.AS20212 cited

Conditional Sound Generation Using Neural Discrete Time-Frequency Representation Learning

Xubo Liu, Turab Iqbal, Jinzheng Zhao +3

Deep generative models have recently achieved impressive performance in speech and music synthesis. However, compared to the generation of those domain-specific sounds, generating…

cs.SD2021

Enhancing Audio Augmentation Methods with Consistency Learning

Turab Iqbal, Karim Helwani, Arvindh Krishnaswamy +1

Data augmentation is an inexpensive way to increase training data diversity and is commonly achieved via transformations of existing data. For tasks such as classification, there i…

cs.SD2020

An Improved Event-Independent Network for Polyphonic Sound Event Localization and Detection

Yin Cao, Turab Iqbal, Qiuqiang Kong +3

Polyphonic sound event localization and detection (SELD), which jointly performs sound event detection (SED) and direction-of-arrival (DoA) estimation, detects the type and occurre…

eess.AS202024 cited

Event-Independent Network for Polyphonic Sound Event Localization and Detection

Yin Cao, Turab Iqbal, Qiuqiang Kong +3

Polyphonic sound event localization and detection is not only detecting what sound events are happening but localizing corresponding sound sources. This series of tasks was first i…

cs.SD20203 cited

Learning with Out-of-Distribution Data for Audio Classification

Turab Iqbal, Yin Cao, Qiuqiang Kong +2

In supervised machine learning, the assumption that training data is labelled correctly is not always satisfied. In this paper, we investigate an instance of labelling error for cl…

cs.SD2019

PANNs: Large-Scale Pretrained Audio Neural Networks for Audio Pattern Recognition

Qiuqiang Kong, Yin Cao, Turab Iqbal +3

Audio pattern recognition is an important research topic in the machine learning area, and includes several tasks such as audio tagging, acoustic scene classification, music classi…