agent training 1large language models 1outcome verification 1reinforcement learning 1self-distillation 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
From Scoring to Acting: Outcome-Verified Comparative Self-Distillation for LLM Agents
Xu Xia, Jinghua Piao, Min Yang +3
The paper introduces Outcome-Verified Comparative Self-Distillation (OVCSD), a method that lets large language model agents internalize skills by supervising them with teachers who…
cs.AI2026
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents
Min Yang, Jinghua Piao, Xu Xia +4
Skills provide an effective mechanism for improving LLM agents on complex tasks, yet in existing agent frameworks, their creation, refinement, and selection are typically governed…
cs.AI2026
ClinicalReTrial: Clinical Trial Redesign with Self-Evolving Agents
Sixue Xing, Kerui Wu, Xuanye Xia +3
Clinical trials constitute a critical yet exceptionally challenging and costly stage of drug development ($2.6B per drug), where protocols are encoded as complex natural language…