4 papers · 1 filter
Challenge on Sound Scene Synthesis: Evaluating Text-to-Audio Generation
Junwon Lee, Modan Tailleur, Laurie M. Heller +5
Despite significant advancements in neural text-to-audio generation, challenges persist in controllability and evaluation. This paper addresses these issues through the Sound Scene…
EMVD dataset: a dataset of extreme vocal distortion techniques used in heavy metal
Modan Tailleur, Julien Pinquier, Laurent Millot +2
In this paper, we introduce the Extreme Metal Vocals Dataset, which comprises a collection of recordings of extreme vocal techniques performed within the realm of heavy metal music…
Detection of Deepfake Environmental Audio
Hafsa Ouajdi, Oussama Hadder, Modan Tailleur +2
With the ever-rising quality of deep generative models, it is increasingly important to be able to discern whether the audio data at hand have been recorded or synthesized. Althoug…
Towards better visualizations of urban sound environments: insights from interviews
Modan Tailleur, Pierre Aumond, Vincent Tourre +1
Urban noise maps and noise visualizations traditionally provide macroscopic representations of noise levels across cities. However, those representations fail at accurately gauging…