22 citations · 27 across the 6 of their papers we have counts for
6 papers
A decade of DCASE: Achievements, practices, evaluations and future challenges
Annamaria Mesaros, Romain Serizel, Toni Heittola +2
This paper introduces briefly the history and growth of the Detection and Classification of Acoustic Scenes and Events (DCASE) challenge, workshop, research area and research commu…
Positive and negative sampling strategies for self-supervised learning on audio-video data
Shanshan Wang, Soumya Tripathy, Toni Heittola +1
In Self-Supervised Learning (SSL), Audio-Visual Correspondence (AVC) is a popular task to learn deep audio and video features from large unlabeled datasets. The key step in AVC is…
Class-Incremental Learning for Multi-Label Audio Classification
Manjunath Mulimani, Annamaria Mesaros
In this paper, we propose a method for class-incremental learning of potentially overlapping sounds for solving a sequence of multi-label audio classification tasks. We design an i…
Evaluating Classification Systems Against Soft Labels with Fuzzy Precision and Recall
Manu Harju, Annamaria Mesaros
Classification systems are normally trained by minimizing the cross-entropy between system outputs and reference labels, which makes the Kullback-Leibler divergence a natural choic…
Training sound event detection with soft labels from crowdsourced annotations
Irene Martín-Morató, Manu Harju, Paul Ahokas +1
In this paper, we study the use of soft labels to train a system for sound event detection (SED). Soft labels can result from annotations which account for human uncertainty about…
Low-complexity acoustic scene classification in DCASE 2022 Challenge
Irene Martín-Morató, Francesco Paissan, Alberto Ancilotto +5
This paper presents an analysis of the Low-Complexity Acoustic Scene Classification task in DCASE 2022 Challenge. The task was a continuation from the previous years, but the low-c…