1 paper
Wuxinlin Cheng, Yupeng Cao, Jinwen Wu +3
Recent strides in pretrained transformer-based language models have propelled state-of-the-art performance in numerous NLP tasks. Yet, as these models grow in size and deployment,…