Showing eess.ASShow all
3 papers · 1 filter
eess.AS2026
WAXAL: A Large-Scale Multilingual African Language Speech Corpus
Abdoulaye Diack, Perry Nelson, Kwaku Agbesi +40
The advancement of speech technology has predominantly favored high-resource languages, creating a significant digital divide for speakers of most Sub-Saharan African languages. To…
eess.AS2025
Towards a Single ASR Model That Generalizes to Disordered Speech
Jimmy Tobin, Katrin Tomanek, Subhashini Venugopalan
This study investigates the impact of integrating a dataset of disordered speech recordings (1,000 hours) into the fine-tuning of a near state-of-the-art ASR baseline system.…
eess.AS2024
Speech Recognition With LLMs Adapted to Disordered Speech Using Reinforcement Learning
Chirag Nagpal, Subhashini Venugopalan, Jimmy Tobin +3
We introduce a large language model (LLM) capable of processing speech inputs and show that tuning it further with reinforcement learning on human preference (RLHF) enables it to a…