3 papers
cs.LG2026
SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History
Zhiwei Li, Yong Hu
Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually change. Existing methods im…
cs.AI2026
MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models
Zhaokang Liao, Yingguo Gao, Yi Yang +2
Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising approach to improve the reasoning abilities of Large Language Models (LLMs). Among RLVR algorithms,…
cs.CL2024
Domain-Specific Improvement on Psychotherapy Chatbot Using Assistant
Cheng Kang, Daniel Novak, Katerina Urbanova +2
Large language models (LLMs) have demonstrated impressive generalization capabilities on specific tasks with human-written instruction data. However, the limited quantity, diversit…