9 citations · 12 across the 2 of their papers we have counts for
1 paper · 1 filter
Chen Liang, Haoming Jiang, Zheng Li +3
Knowledge distillation has been shown to be a powerful model compression approach to facilitate the deployment of pre-trained language models in practice. This paper focuses on tas…