2 papers
cs.CL2026
Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild
Berkin Durmus, Chen Cen, Eduardo Pacheco +2
The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes domains suggest otherwise. We…
cs.SD2025
WhisperKit: On-device Real-time ASR with Billion-Scale Transformers
Atila Orhon, Arda Okan, Berkin Durmus +2
Real-time Automatic Speech Recognition (ASR) is a fundamental building block for many commercial applications of ML, including live captioning, dictation, meeting transcriptions, a…