168 citations · 178 across the 10 of their papers we have counts for
4 papers · 1 filter
Mitigating Language-Level Performance Disparity in mPLMs via Teacher Language Selection and Cross-lingual Self-Distillation
Haozhe Zhao, Zefan Cai, Shuzheng Si +4
Large-scale multilingual Pretrained Language Models (mPLMs) yield impressive performance on cross-language tasks, yet significant performance disparities exist across different lan…
Guiding AMR Parsing with Reverse Graph Linearization
Bofei Gao, Liang Chen, Peiyi Wang +2
Abstract Meaning Representation (AMR) parsing aims to extract an abstract semantic graph from a given sentence. The sequence-to-sequence approaches, which linearize the semantic gr…
HomoDistil: Homotopic Task-Agnostic Distillation of Pre-trained Transformers
Chen Liang, Haoming Jiang, Zheng Li +3
Knowledge distillation has been shown to be a powerful model compression approach to facilitate the deployment of pre-trained language models in practice. This paper focuses on tas…
Definition Modeling: Learning to define word embeddings in natural language
Thanapon Noraset, Chen Liang, Larry Birnbaum +1
Distributed representations of words have been shown to capture lexical semantics, as demonstrated by their effectiveness in word similarity and analogical relation tasks. But, the…