Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Coalition-Aware Skill Reliability for Self-Evolving Agents
Qiyan Zhao, Xiaofeng Zhang, Bo Liu +11
Agent skills, structured artifacts distilled from interaction trajectories and dynamically reused from skill banks, have become a central mechanism for enabling large language mode…
cs.AI2025
SDPO: Segment-Level Direct Preference Optimization for Social Agents
Aobo Kong, Wentao Ma, Shiwan Zhao +7
Social agents powered by large language models (LLMs) can simulate human social behaviors but fall short in handling complex social dialogues. Direct Preference Optimization (DPO)…