2 papers
cs.CL2023
A Comparative Analysis of Task-Agnostic Distillation Methods for Compressing Transformer Language Models
Takuma Udagawa, Aashka Trivedi, Michele Merler +1
Large language models have become a vital component in modern NLP, achieving state of the art performance in a variety of tasks. However, they are often inefficient for real-world…
cs.CL2023
CoSiNES: Contrastive Siamese Network for Entity Standardization
Jiaqing Yuan, Michele Merler, Mihir Choudhury +3
Entity standardization maps noisy mentions from free-form text to standard entities in a knowledge base. The unique challenge of this task relative to other entity-related tasks is…