1 paper
Hwanjun Song, Igor Shalyminov, Hang Su +3
Sequence-level knowledge distillation reduces the size of Seq2Seq models for more efficient abstractive summarization. However, it often leads to a loss of abstractiveness in summa…