2 papers
cs.CL2025
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
Michael Ong, Sean Robertson, Leo Peckham +7
We introduce the Faetar Automatic Speech Recognition Benchmark, a benchmark corpus designed to push the limits of current approaches to low-resource speech recognition. Faetar, a F…
cs.CL2024
Quantifying the Role of Textual Predictability in Automatic Speech Recognition
Sean Robertson, Gerald Penn, Ewan Dunbar
A long-standing question in automatic speech recognition research is how to attribute errors to the ability of a model to model the acoustics, versus its ability to leverage higher…