Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Peeking Into The Future For Contextual Biasing
Ramaneswaran Selvakumar, Cindy Tseng, Eesung Kim +2
While end-to-end (E2E) automatic speech recognition (ASR) models excel at general transcription, they struggle to recognize rare or unseen named entities (e.g., contact names, loca…
cs.CL2025
Enhanced Hybrid Transducer and Attention Encoder Decoder with Text Data
Yun Tang, Eesung Kim, Vijendra Raj Apsingekar
A joint speech and text optimization method is proposed for hybrid transducer and attention-based encoder decoder (TAED) modeling to leverage large amounts of text corpus and enhan…
cs.CL2024
Transducer Consistency Regularization for Speech to Text Applications
Cindy Tseng, Yun Tang, Vijendra Raj Apsingekar
Consistency regularization is a commonly used practice to encourage the model to generate consistent representation from distorted input features and improve model generalization.…