1 paper
Changyu Hou, Jun Wang, Yixuan Qiao +8
Large scale pre-training models have been widely used in named entity recognition (NER) tasks. However, model ensemble through parameter averaging or voting can not give full play…