1 paper
Shuo Xie, Jiahao Qiu, Ankita Pasad +3
While transferring a pretrained language model, common approaches conventionally attach their task-specific classifiers to the top layer and adapt all the pretrained layers. We inv…