1 citations · 2 across the 17 of their papers we have counts for
12 papers · 1 filter
ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains
Ziqi Zhao, Xinyu Ma, Liu Yang +6
On-policy self-distillation (OPSD) improves the reasoning performance of large language models (LLMs) by providing dense token-level supervision for on-policy rollouts. However, ex…
Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models
Yujie Feng, Jian Li, Zhihan Zhou +7
Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation where redundant retrieved contex…
AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning
Yujie Feng, Jian Li, Xiaoyu Dong +8
Continual learning (CL) is essential for deploying large language models (LLMs) in dynamic real-world environments without the need for costly retraining. Recent model merging-base…
REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering
Li-Ming Zhan, Bo Liu, Chengqiang Xie +2
Inference-time steering aims to alter a large language model's (LLM's) responses without changing its parameters, but a central challenge is identifying the internal modules that m…
GeoEdit: Geometric Knowledge Editing for Large Language Models
Yujie Feng, Liming Zhan, Zexin Lu +6
Regular updates are essential for maintaining up-to-date knowledge in large language models (LLMs). Consequently, various model editing methods have been developed to update specif…
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
Bo Liu, Liming Zhan, Yujie Feng +5
In the realm of task-oriented dialogue systems, a robust intent detection mechanism must effectively handle malformed utterances encountered in real-world scenarios. This study pre…