3 citations
1 paper
Ta-Chun Su, Hsiang-Chih Cheng
Fine-tuning with pre-trained models has achieved exceptional results for many language tasks. In this study, we focused on one such self-attention network model, namely BERT, which…