Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
Towards Quantifying Benchmark Optimization in ASR Models
Theo Lebryk, David Ayllon, Alice Baird +3
Public benchmarks are important measures of Automatic Speech Recognition (ASR) model capabilities. However, by nature of being public, there is risk of models being optimized for t…
cs.SD2026
RW-Voice-EQ Bench: A Real World Benchmark for Evaluating Voice AI Systems
David Ayllon, Alice Baird, Jeffrey Brooks +11
Current voice AI benchmarks typically evaluate isolated capabilities such as speech intelligibility, word error rate, or text-based dialogue quality, but they rarely test whether s…