Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
Yanfan Du, Jun Zhang, Bin Wang +6
Recent advances in speech large language models (SLMs) have improved speech recognition and translation in general domains, but accurately generating domain-specific terms or neolo…
cs.CL2023
Improving Large-scale Deep Biasing with Phoneme Features and Text-only Data in Streaming Transducer
Jin Qiu, Lu Huang, Boyu Li +3
Deep biasing for the Transducer can improve the recognition performance of rare words or contextual entities, which is essential in practical applications, especially for streaming…
cs.CL2023
Text-only Domain Adaptation using Unified Speech-Text Representation in Transducer
Lu Huang, Boyu Li, Jun Zhang +2
Domain adaptation using text-only corpus is challenging in end-to-end(E2E) speech recognition. Adaptation by synthesizing audio from text through TTS is resource-consuming. We pres…