3 papers
Fine-tuning Pre-trained Audio Models for COVID-19 Detection: A Technical Report
Daniel Oliveira de Brito, Letícia Gabriella de Souza, Marcelo Matheus Gauy +2
This technical report investigates the performance of pre-trained audio models on COVID-19 detection tasks using established benchmark datasets. We fine-tuned Audio-MAE and three P…
The Impact of Prosodic Segmentation on Speech Synthesis of Spontaneous Speech
Julio Cesar Galdino, Sidney Evaldo Leal, Leticia Gabriella De Souza +6
Spontaneous speech presents several challenges for speech synthesis, particularly in capturing the natural flow of conversation, including turn-taking, pauses, and disfluencies. Al…
Bringing NURC/SP to Digital Life: the Role of Open-source Automatic Speech Recognition Models
Lucas Rafael Stefanel Gris, Arnaldo Candido Junior, Vinícius G. dos Santos +4
The NURC Project that started in 1969 to study the cultured linguistic urban norm spoken in five Brazilian capitals, was responsible for compiling a large corpus for each capital.…