30 citations · 48 across the 5 of their papers we have counts for
5 papers · 1 filter
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
Riccardo Passoni, Francesca Ronchini, Luca Comanducci +2
Text-to-audio models have recently emerged as a powerful technology for generating sound from textual descriptions. However, their high computational demands raise concerns about e…
DCASE 2024 Task 4: Sound Event Detection with Heterogeneous Data and Missing Labels
Samuele Cornell, Janek Ebbers, Constance Douwes +4
The Detection and Classification of Acoustic Scenes and Events Challenge Task 4 aims to advance sound event detection (SED) systems in domestic environments by leveraging training…
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
Francesca Ronchini, Romain Serizel
In recent years, deep learning systems have shown a concerning trend toward increased complexity and higher energy consumption. As researchers in this domain and organizers of one…
Improving Sound Event Detection Metrics: Insights from DCASE 2020
Giacomo Ferroni, Nicolas Turpault, Juan Azcarreta +4
The ranking of sound event detection (SED) systems may be biased by assumptions inherent to evaluation criteria and to the choice of an operating point. This paper compares convent…
The Speed Submission to DIHARD II: Contributions & Lessons Learned
Md Sahidullah, Jose Patino, Samuele Cornell +11
This paper describes the speaker diarization systems developed for the Second DIHARD Speech Diarization Challenge (DIHARD II) by the Speed team. Besides describing the system, whic…