2 papers
cs.CL2026
VocalAffectBench: Evaluating Vocal Emotion Recognition in AI Audio Models
Luc Debaupte, Tyler Baumgartner, Brandon Tai +3
Voice products increasingly need affective cues that are present in speech but absent from transcripts. We introduce VocalAffectBench, a public, test-only benchmark for evaluating…
cs.CL2026
VoiceCodeBench: Evaluating Exact Structured-Token Recovery in Automatic Speech Recognition
Tyler Baumgartner, Brandon Tai, Lisa Kaelin-Martin +4
Automatic speech recognition (ASR) systems are commonly evaluated with word error rate (WER), yet many voice workflows depend on exact written values for identifiers, paths, and me…