15 citations · 35 across the 12 of their papers we have counts for
3 papers · 1 filter
Phonemic Representation and Transcription for Speech to Text Applications for Under-resourced Indigenous African Languages: The Case of Kiswahili
Ebbie Awino, Lilian Wanzare, Lawrence Muchemi +5
Building automatic speech recognition (ASR) systems is a challenging task, especially for under-resourced languages that need to construct corpora nearly from scratch and lack suff…
Kencorpus: A Kenyan Language Corpus of Swahili, Dholuo and Luhya for Natural Language Processing Tasks
Barack Wanjawa, Lilian Wanzare, Florence Indede +3
Indigenous African languages are categorized as under-served in Natural Language Processing. They therefore experience poor digital inclusivity and information access. The processi…
KenSwQuAD -- A Question Answering Dataset for Swahili Low Resource Language
Barack W. Wanjawa, Lilian D. A. Wanzare, Florence Indede +3
The need for Question Answering datasets in low resource languages is the motivation of this research, leading to the development of Kencorpus Swahili Question Answering Dataset, K…