Integrating the Data Augmentation Scheme with Various Classifiers for Acoustic Scene Modeling
arXiv:1907.06639
Abstract
This technical report describes the IOA team's submission for TASK1A of DCASE2019 challenge. Our acoustic scene classification (ASC) system adopts a data augmentation scheme employing generative adversary networks. Two major classifiers, 1D deep convolutional neural network integrated with scalogram features and 2D fully convolutional neural network integrated with Mel filter bank features, are deployed in the scheme. Other approaches, such as adversary city adaptation, temporal module based on discrete cosine transform and hybrid architectures, have been developed for further fusion. The results of our experiments indicates that the final fusion systems A-D could achieve an accuracy higher than 85% on the officially provided fold 1 evaluation dataset.
Cited by in corpus (7)
- Receptive Field Regularization Techniques for Audio Classification and Tagging with Deep Convolutional Neural Networks
- Acoustic Scene Classification with Spectrogram Processing Strategies
- Low-Complexity Models for Acoustic Scene Classification Based on Receptive Field Regularization and Frequency Damping
- Relational Teacher Student Learning with Neural Label Embedding for Device Adaptation in Acoustic Scene Classification
- A Curated Dataset of Urban Scenes for Audio-Visual Scene Analysis
- DD-CNN: Depthwise Disout Convolutional Neural Network for Low-complexity Acoustic Scene Classification
- An Acoustic Segment Model Based Segment Unit Selection Approach to Acoustic Scene Classification with Partial Utterances