3 citations · 3 across the 2 of their papers we have counts for
6 papers
What is the ground truth? Reliability of multi-annotator data for audio tagging
Irene Martin-Morato, Annamaria Mesaros
Crowdsourcing has become a common approach for annotating large amounts of data. It has the advantage of harnessing a large workforce to produce large amounts of data in a short ti…
A Curated Dataset of Urban Scenes for Audio-Visual Scene Analysis
Shanshan Wang, Annamaria Mesaros, Toni Heittola +1
This paper introduces a curated dataset of urban scenes for audio-visual scene analysis which consists of carefully selected and recorded material. The data was recorded in multipl…
Overview and Evaluation of Sound Event Localization and Detection in DCASE 2019
Archontis Politis, Annamaria Mesaros, Sharath Adavanne +2
Sound event localization and detection is a novel area of research that emerged from the combined interest of analyzing the acoustic scene in terms of the spatial and temporal acti…
Acoustic scene classification in DCASE 2020 Challenge: generalization across devices and low complexity solutions
Toni Heittola, Annamaria Mesaros, Tuomas Virtanen
This paper presents the details of Task 1: Acoustic Scene Classification in the DCASE 2020 Challenge. The task consists of two subtasks: classification of data from multiple device…
City classification from multiple real-world sound scenes
Helen L. Bear, Toni Heittola, Annamaria Mesaros +2
The majority of sound scene analysis work focuses on one of two clearly defined tasks: acoustic scene classification or sound event detection. Whilst this separation of tasks is us…
A multi-device dataset for urban acoustic scene classification
Annamaria Mesaros, Toni Heittola, Tuomas Virtanen
This paper introduces the acoustic scene classification task of DCASE 2018 Challenge and the TUT Urban Acoustic Scenes 2018 dataset provided for the task, and evaluates the perform…