4 papers
SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History
Zhiwei Li, Yong Hu
Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually change. Existing methods im…
MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models
Zhaokang Liao, Yingguo Gao, Yi Yang +2
Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising approach to improve the reasoning abilities of Large Language Models (LLMs). Among RLVR algorithms,…
Domain-Specific Improvement on Psychotherapy Chatbot Using Assistant
Cheng Kang, Daniel Novak, Katerina Urbanova +2
Large language models (LLMs) have demonstrated impressive generalization capabilities on specific tasks with human-written instruction data. However, the limited quantity, diversit…
Quantized Embedding Vectors for Controllable Diffusion Language Models
Cheng Kang, Xinye Chen, Yong Hu +1
Improving the controllability, portability, and inference speed of diffusion language models (DLMs) is a key challenge in natural language generation. While recent research has sho…