3 citations · 3 across the 1 of their papers we have counts for
1 paper
Ta-Chun Su, Hsiang-Chih Cheng
Fine-tuning with pre-trained models has achieved exceptional results for many language tasks. In this study, we focused on one such self-attention network model, namely BERT, which…