1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2023
Human Transcription Quality Improvement
Jian Gao, Hanbo Sun, Cheng Cao +1
High quality transcription data is crucial for training automatic speech recognition (ASR) systems. However, the existing industry-level data collection pipelines are expensive to…
eess.AS2023
HTEC: Human Transcription Error Correction
Hanbo Sun, Jian Gao, Xiaomin Wu +3
High-quality human transcription is essential for training and improving Automatic Speech Recognition (ASR) models. Recent study~\cite{libricrowd} has found that every 1% worse tra…
cs.CL2023★ 1 cited
APAM: Adaptive Pre-training and Adaptive Meta Learning in Language Model for Noisy Labels and Long-tailed Learning
Sunyi Chi, Bo Dong, Yiming Xu +2
Practical natural language processing (NLP) tasks are commonly long-tailed with noisy labels. Those problems challenge the generalization and robustness of complex models such as D…