3 papers
cs.CL2026
Earnings25: A Comprehensive 500-Hour Speech Benchmark for Finance
Denglin Jiang, Haoran Zhou, Anshul Wadhawan +7
We introduce Earnings25, a finance-domain benchmark for evaluating automatic speech recognition (ASR) on English-language earnings calls under realistic conditions. Earnings25 comp…
cs.AI2026
PExA: Parallel Exploration Agent for Complex Text-to-SQL
Tanmay Parekh, Ella Hofmann-Coyle, Shuyi Wang +3
LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice versa. We reformulate text-to…
cs.SD2025
Adapting Whisper for Streaming Speech Recognition via Two-Pass Decoding
Haoran Zhou, Xingchen Song, Brendan Fahy +9
OpenAI Whisper is a family of robust Automatic Speech Recognition (ASR) models trained on 680,000 hours of audio. However, its encoder-decoder architecture, trained with a sequence…