4 papers
WAXAL: A Large-Scale Multilingual African Language Speech Corpus
Abdoulaye Diack, Perry Nelson, Kwaku Agbesi +40
The advancement of speech technology has predominantly favored high-resource languages, creating a significant digital divide for speakers of most Sub-Saharan African languages. To…
Enabling Automatic Disordered Speech Recognition: An Impaired Speech Dataset in the Akan Language
Isaac Wiafe, Akon Obu Ekpezu, Sumaya Ahmed Salihs +3
The lack of impaired speech data hinders advancements in the development of inclusive speech technologies, particularly in low-resource languages such as Akan. To address this gap,…
A Cookbook for Community-driven Data Collection of Impaired Speech in LowResource Languages
Sumaya Ahmed Salihs, Isaac Wiafe, Jamal-Deen Abdulai +7
This study presents an approach for collecting speech samples to build Automatic Speech Recognition (ASR) models for impaired speech, particularly, low-resource languages. It aims…
Benchmarking Akan ASR Models Across Domain-Specific Datasets: A Comparative Evaluation of Performance, Scalability, and Adaptability
Mark Atta Mensah, Isaac Wiafe, Akon Ekpezu +5
Most existing automatic speech recognition (ASR) research evaluate models using in-domain datasets. However, they seldom evaluate how they generalize across diverse speech contexts…