3 citations
1 paper · 1 filter
Ta-Chun Su, Hsiang-Chih Cheng
Fine-tuning with pre-trained models has achieved exceptional results for many language tasks. In this study, we focused on one such self-attention network model, namely BERT, which…