2 papers
cs.CL2025
Peeking Into The Future For Contextual Biasing
Ramaneswaran Selvakumar, Cindy Tseng, Eesung Kim +2
While end-to-end (E2E) automatic speech recognition (ASR) models excel at general transcription, they struggle to recognize rare or unseen named entities (e.g., contact names, loca…
cs.CL2025
Enhanced Hybrid Transducer and Attention Encoder Decoder with Text Data
Yun Tang, Eesung Kim, Vijendra Raj Apsingekar
A joint speech and text optimization method is proposed for hybrid transducer and attention-based encoder decoder (TAED) modeling to leverage large amounts of text corpus and enhan…