2 papers
cs.CL2026
RAISE: Reinforced Adaptive Instruction Selection For Large Language Models
Qingsong Lv, Yangning Li, Zihua Lan +8
In the instruction fine-tuning of large language models (LLMs), it is widely recognized that a few high-quality instructions are superior to a large number of low-quality instructi…
cs.CL2025
MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
Yangning Li, Zihua Lan, Lv Qingsong +2
As Large Language Models (LLMs) are increasingly applied across various tasks, instruction tuning has emerged as a critical method for enhancing model performance. However, current…