2 papers
cs.AI2026
SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior
Zhiyu Chen, Zihan Guo, Bo Huang +4
Agent Skills augment large language model (LLM) agents with procedural knowledge at inference time, but current benchmarks rarely distinguish what a Skill says from how it is organ…
cs.MA2026
MonoScale: Scaling Multi-Agent System with Monotonic Improvement
Shuai Shao, Yixiang Liu, Bingwei Lu +1
In recent years, LLM-based multi-agent systems (MAS) have advanced rapidly, using a router to decompose tasks and delegate subtasks to specialized agents. A natural way to expand c…