Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO
Yu Tian, Jiawei Chen, Lifan Zheng +5
We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-based agents. Addressing the…
cs.CL2026
LCO: LLM-based Constraint Optimization for Safer Agentic LLMs in Real-world Tasks
Jiayong Wan, Jiawei Chen, Zhaoxia Yin +2
Large Language Models (LLMs) are increasingly acting as autonomous agents, but their continuous interaction with the environment can lead to in-context reward hacking (ICRH), a phe…