3 papers
cs.CL2026
Continue, Adapt, or Yield: In-Turn Adaptation to Overlapping Speech in Full-Duplex Agents
Yunqi Lu, Tyler Baumgartner, Nikhil Johri +6
Full-duplex evaluation often emphasizes whether an agent keeps speaking or stops. That binary cannot express a third response humans use routinely: continuing to speak while incorp…
cs.CL2026
VocalAffectBench: Evaluating Vocal Emotion Recognition in AI Audio Models
Luc Debaupte, Tyler Baumgartner, Brandon Tai +3
Voice products increasingly need affective cues that are present in speech but absent from transcripts. We introduce VocalAffectBench, a public, test-only benchmark for evaluating…
cs.CL2026
VoiceCodeBench: Evaluating Exact Structured-Token Recovery in Automatic Speech Recognition
Tyler Baumgartner, Brandon Tai, Lisa Kaelin-Martin +4
Automatic speech recognition (ASR) systems are commonly evaluated with word error rate (WER), yet many voice workflows depend on exact written values for identifiers, paths, and me…