31 citations · 92 across the 20 of their papers we have counts for
15 papers · 1 filter
Making Pre-trained Language Models both Task-solvers and Self-calibrators
Yangyi Chen, Xingyao Wang, Heng Ji
Pre-trained language models (PLMs) serve as backbones for various real-world systems. For high-stake applications, it's equally essential to have reasonable confidence estimations…
OpenPI-C: A Better Benchmark and Stronger Baseline for Open-Vocabulary State Tracking
Xueqing Wu, Sha Li, Heng Ji
Open-vocabulary state tracking is a more practical version of state tracking that aims to track state changes of entities throughout a process without restricting the state space a…
C-PMI: Conditional Pointwise Mutual Information for Turn-level Dialogue Evaluation
Liliang Ren, Mankeerat Sidhu, Qi Zeng +3
Existing reference-free turn-level evaluation metrics for chatbots inadequately capture the interaction between the user and the system. Consequently, they often correlate poorly w…
Understanding the Effect of Data Augmentation on Knowledge Distillation
Ziqi Wang, Chi Han, Wenxuan Bao +1
Knowledge distillation (KD) requires sufficient data to transfer knowledge from large-scale teacher models to small-scale student models. Therefore, data augmentation has been wide…
Information Association for Language Model Updating by Mitigating LM-Logical Discrepancy
Pengfei Yu, Heng Ji
Large Language Models~(LLMs) struggle with providing current information due to the outdated pre-training data. Existing methods for updating LLMs, such as knowledge editing and co…
CREATOR: Tool Creation for Disentangling Abstract and Concrete Reasoning of Large Language Models
Cheng Qian, Chi Han, Yi R. Fung +3
Large Language Models (LLMs) have made significant progress in utilizing tools, but their ability is limited by API availability and the instability of implicit reasoning, particul…