18 citations · 18 across the 2 of their papers we have counts for
2 papers
cs.CL2022
Language Model Pre-Training with Sparse Latent Typing
Liliang Ren, Zixuan Zhang, Han Wang +3
Modern large-scale Pre-trained Language Models (PLMs) have achieved tremendous success on a wide range of downstream tasks. However, most of the LM pre-training objectives only foc…
cs.CL2014★ 18 cited
Scalable Topical Phrase Mining from Text Corpora
Ahmed El-Kishky, Yanglei Song, Chi Wang +2
While most topic modeling algorithms model text corpora with unigrams, human interpretation often relies on inherent grouping of terms into phrases. As such, we consider the proble…