23 citations · 23 across the 2 of their papers we have counts for
2 papers
cs.CL2022
Improving Contextual Representation with Gloss Regularized Pre-training
Yu Lin, Zhecheng An, Peihao Wu +1
Though achieving impressive results on many NLP tasks, the BERT-like masked language models (MLM) encounter the discrepancy between pre-training and inference. In light of this gap…
cs.CL2017★ 23 cited
Deep LSTM for Large Vocabulary Continuous Speech Recognition
Xu Tian, Jun Zhang, Zejun Ma +6
Recurrent neural networks (RNNs), especially long short-term memory (LSTM) RNNs, are effective network for sequential task like speech recognition. Deeper LSTM models perform well…