activity
20242026
collaborators

9 papers

cs.CL2026

ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains

Ziqi Zhao, Xinyu Ma, Liu Yang +6

On-policy self-distillation (OPSD) improves the reasoning performance of large language models (LLMs) by providing dense token-level supervision for on-policy rollouts. However, ex…

cs.LG2026

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

Yujie Feng, Hao Wang, Jian Li +6

Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memory replay methods are widely used for…

cs.CL2026

Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models

Yujie Feng, Jian Li, Zhihan Zhou +7

Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation where redundant retrieved contex…

cs.CL2025

AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning

Yujie Feng, Jian Li, Xiaoyu Dong +8

Continual learning (CL) is essential for deploying large language models (LLMs) in dynamic real-world environments without the need for costly retraining. Recent model merging-base…

cs.LG2025

Recurrent Knowledge Identification and Fusion for Language Model Continual Learning

Yujie Feng, Xujia Wang, Zexin Lu +7

Continual learning (CL) is crucial for deploying large language models (LLMs) in dynamic real-world environments without costly retraining. While recent model ensemble and model me…

cs.CL2025

GeoEdit: Geometric Knowledge Editing for Large Language Models

Yujie Feng, Liming Zhan, Zexin Lu +6

Regular updates are essential for maintaining up-to-date knowledge in large language models (LLMs). Consequently, various model editing methods have been developed to update specif…