35 citations · 65 across the 3 of their papers we have counts for
3 papers
cs.CL2022★ 13 cited
Momentum Calibration for Text Generation
Xingxing Zhang, Yiran Liu, Xun Wang +5
The input and output of most text generation tasks can be transformed to two sequences of tokens and they can be modeled using sequence-to-sequence learning modeling tools such as…
eess.AS2019★ 17 cited
Advances in Online Audio-Visual Meeting Transcription
Takuya Yoshioka, Igor Abramovski, Cem Aksoylar +23
This paper describes a system that generates speaker-annotated transcripts of meetings by using a microphone array and a 360-degree camera. The hallmark of the system is its abilit…
cs.CL2017★ 35 cited
Progressive Joint Modeling in Unsupervised Single-channel Overlapped Speech Recognition
Zhehuai Chen, Jasha Droppo, Jinyu Li +1
Unsupervised single-channel overlapped speech recognition is one of the hardest problems in automatic speech recognition (ASR). Permutation invariant training (PIT) is a state of t…