116 citations · 316 across the 17 of their papers we have counts for
1 paper · 2 filters
Yiren Chen, Xiaoyu Kou, Jiangang Bai +1
One of the most popular paradigms of applying large pre-trained NLP models such as BERT is to fine-tune it on a smaller dataset. However, one challenge remains as the fine-tuned mo…