Publications (14)
MambaFoley: Foley Sound Generation using Selective State-Space Models
Marco Furio Colombo, Francesca Ronchini, Luca Comanducci +1
Recent advancements in deep learning have led to widespread use of techniques for audio content generation, notably employing Denoising Diffusion Probabilistic Models (DDPM) across…
Sound event localization and detection based on crnn using rectangular filters and channel rotation data augmentation
Francesca Ronchini, Daniel Arteaga, Andrés Pérez-López
Sound Event Localization and Detection refers to the problem of identifying the presence of independent or temporally-overlapped sound sources, correctly identifying to which sound…
The impact of non-target events in synthetic soundscapes for sound event detection
Francesca Ronchini, Romain Serizel, Nicolas Turpault +1
Detection and Classification Acoustic Scene and Events Challenge 2021 Task 4 uses a heterogeneous dataset that includes both recorded and synthetic soundscapes. Until recently only…
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
Francesca Ronchini, Romain Serizel
This paper proposes a benchmark of submissions to Detection and Classification Acoustic Scene and Events 2021 Challenge (DCASE) Task 4 representing a sampling of the state-of-the-a…
Synthetic training set generation using text-to-audio models for environmental sound classification
Francesca Ronchini, Luca Comanducci, Fabio Antonacci
In recent years, text-to-audio models have revolutionized the field of automatic audio generation. This paper investigates their application in generating synthetic datasets for tr…
PAGURI: a user experience study of creative interaction with text-to-music models
Francesca Ronchini, Luca Comanducci, Gabriele Perego +1
In recent years, text-to-music models have been the biggest breakthrough in automatic music generation. While they are unquestionably a showcase of technological progress, it is no…