2 papers
cs.AI2026
Diagnosing with Insights: Structured Analysis of Agent Failures via Behavioral Abstractions
Jiayi Bi, Yanjie Gao, Yuanmin Xie +4
With the proliferation of LLM agents, the ability to understand and diagnose failures in agents is essential to achieving superior effectiveness and trustworthiness. As agent failu…
cs.CR2026
ShadowProbe: Language-Extensible Detection of Hidden Algorithmic Complexity Vulnerabilities
Yuanmin Xie, Xiangfan Wu, Wenhao Wu +6
Algorithmic Complexity Vulnerabilities (ACVs) arise when adversarial inputs trigger worst-case execution behavior, causing severe performance degradation or Denial-of-Service condi…