3 papers
cs.MA2026
Self-Evolving Multi-Agent Systems via Decentralized Memory
Guangya Hao, Yunbo Long, Zhuokai Zhao
Self-evolving multi-agent systems (MAS) have emerged as a promising route to LLM agents that continually improve from experience, with persistent memory at their foundation. Howeve…
cs.CL2026
Self-Policy Distillation via Capability-Selective Subspace Projection
Guangya Hao, Yitong Shang, Yunbo Long +2
Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signals to curate self-generated o…
cs.LG2026
Self-Improving Tabular Language Models via Iterative Reward-Guided Post-Training
Yunbo Long, Tejumade Afonja, Guangya Hao +2
Tabular language models can generate synthetic tables by modeling rows as token sequences, but they are typically trained once with supervised fine-tuning and then used as static s…