2 papers
cs.AI2026
Localizing Emergent Failures in Agentic AI: Recovering Minimal Repair Families via Counterfactual Replay
Bingjie Li, Yumeng Song, Zhongming Yao +1
Failures in agentic AI systems can arise from interactions among messages exchanged by multiple large language model (LLM) agents. Pointwise attribution cannot distinguish a jointl…
cs.CR2026
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
Zonghao Ying, Xiao Yang, Siyang Wu +7
The rapid evolution of Large Language Models (LLMs) into autonomous, tool-calling agents has fundamentally altered the cybersecurity landscape. Frameworks like OpenClaw grant AI sy…