3 papers
cs.LG2026
SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment
Chaofan Meng, Yuhang Zheng, Yingnan Zhou +1
Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill selection. Recent Agent Skill r…
cs.AI2026
Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +306
Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deployment across many professional d…
cs.SE2024
Rethinking Software Misconfigurations in the Real World: An Empirical Study and Literature Analysis
Yuhao Liu, Yingnan Zhou, Hanfeng Zhang +6
Software misconfiguration has consistently been a major reason for software failures. Over the past two decades, much work has been done to detect and diagnose software misconfigur…