4 papers
Listen, Attend, Understand: a Regularization Technique for Stable E2E Speech Translation Training on High Variance labels
Yacouba Diarra, Michael Leventhal
End-to-End Speech Translation often shows slower convergence and worse performance when target transcriptions exhibit high variance and semantic ambiguity. We propose Listen, Atten…
Kunnafonidilaw ka Cadeau: an ASR dataset of present-day Bambara
Yacouba Diarra, Panga Azazia Kamate, Nouhoum Souleymane Coulibaly +1
We present Kunkado, a 160-hour Bambara ASR dataset compiled from Malian radio archives to capture present-day spontaneous speech across a wide range of topics. It includes code-swi…
Dealing with the Hard Facts of Low-Resource African NLP
Yacouba Diarra, Nouhoum Souleymane Coulibaly, Panga Azazia Kamaté +4
Creating speech datasets, models, and evaluation frameworks for low-resource languages remains challenging given the lack of a broad base of pertinent experience to draw from. This…
Cost Analysis of Human-corrected Transcription for Predominately Oral Languages
Yacouba Diarra, Nouhoum Souleymane Coulibaly, Michael Leventhal
Creating speech datasets for low-resource languages is a critical yet poorly understood challenge, particularly regarding the actual cost in human labor. This paper investigates th…