4 papers · 1 filter
DoCIA: An Online Document-Level Context Incorporation Agent for Speech Translation
Xinglin Lyu, Wei Tang, Yuang Li +7
Document-level context is crucial for handling discourse challenges in text-to-text document-level machine translation (MT). Despite the increased discourse challenges introduced b…
Adapting Large Language Models to Log Analysis with Interpretable Domain Knowledge
Yuhe Ji, Yilun Liu, Feiyu Yao +10
Log analysis represents a critical sub-domain within AI applications that facilitates automatic approaches to fault and error management of large-scaled software systems, saving la…
Hard-Synth: Synthesizing Diverse Hard Samples for ASR using Zero-Shot TTS and LLM
Jiawei Yu, Yuang Li, Xiaosong Qiao +6
Text-to-speech (TTS) models have been widely adopted to enhance automatic speech recognition (ASR) systems using text-only corpora, thereby reducing the cost of labeling real speec…
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction
Yuang Li, Xiaosong Qiao, Xiaofeng Zhao +4
Large language models can enhance automatic speech recognition systems through generative error correction. In this paper, we propose Pinyin-enhanced GEC, which leverages Pinyi, th…