5 papers
Kakugo: Distillation of Low-Resource Languages into Small Language Models
Peter Devine, Mardhiyah Sanni, Farid Adilazuarda +2
We present Kakugo, a novel and cost-effective pipeline designed to train general-purpose Small Language Models (SLMs) for low-resource languages using only the language name as inp…
AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR
Gabrial Zencha Ashungafac, Mardhiyah Sanni, Busayo Awobade +2
Recent advances in speech-enabled AI, including Google's NotebookLM and OpenAI's speech-to-speech API, are driving widespread interest in voice interfaces globally. Despite this mo…
AfriMed-QA: A Pan-African, Multi-Specialty, Medical Question-Answering Benchmark Dataset
Tobi Olatunji, Charles Nimo, Abraham Owodunni +23
Recent advancements in large language model(LLM) performance on medical multiple choice question (MCQ) benchmarks have stimulated interest from healthcare providers and patients gl…
Afrispeech-Dialog: A Benchmark Dataset for Spontaneous English Conversations in Healthcare and Beyond
Mardhiyah Sanni, Tassallah Abdullahi, Devendra D. Kayande +9
Speech technologies are transforming interactions across various sectors, from healthcare to call centers and robots, yet their performance on African-accented conversations remain…
The Multicultural Medical Assistant: Can LLMs Improve Medical ASR Errors Across Borders?
Ayo Adedeji, Mardhiyah Sanni, Emmanuel Ayodele +2
The global adoption of Large Language Models (LLMs) in healthcare shows promise to enhance clinical workflows and improve patient outcomes. However, Automatic Speech Recognition (A…