66 citations · 101 across the 26 of their papers we have counts for
Showing 2022 · cs.CLShow all
2 papers · 2 filters
cs.CL2022
Unsupervised Boundary-Aware Language Model Pretraining for Chinese Sequence Labeling
Peijie Jiang, Dingkun Long, Yanzhao Zhang +3
Boundary information is critical for various Chinese language processing tasks, such as word segmentation, part-of-speech tagging, and named entity recognition. Previous studies us…
cs.CL2022★ 2 cited
Retrieval Oriented Masking Pre-training Language Model for Dense Passage Retrieval
Dingkun Long, Yanzhao Zhang, Guangwei Xu +1
Pre-trained language model (PTM) has been shown to yield powerful text representations for dense passage retrieval task. The Masked Language Modeling (MLM) is a major sub-task of t…