3 papers
cs.CL2025
Chunk Based Speech Pre-training with High Resolution Finite Scalar Quantization
Yun Tang, Cindy Tseng
Low latency speech human-machine communication is becoming increasingly necessary as speech technology advances quickly in the last decade. One of the primary factors behind the ad…
cs.CL2025
Peeking Into The Future For Contextual Biasing
Ramaneswaran Selvakumar, Cindy Tseng, Eesung Kim +2
While end-to-end (E2E) automatic speech recognition (ASR) models excel at general transcription, they struggle to recognize rare or unseen named entities (e.g., contact names, loca…
cs.CL2024
Transducer Consistency Regularization for Speech to Text Applications
Cindy Tseng, Yun Tang, Vijendra Raj Apsingekar
Consistency regularization is a commonly used practice to encourage the model to generate consistent representation from distorted input features and improve model generalization.…