From the 1 of 2 linked papers with an AI index.
2 papers
cs.CR2026
Phantom Guardrails: When Self-Improving Agent Harnesses Fix Failures That Never Happened
Su Wang, Pin Qian, Yifan Lin +5
The paper investigates how self‑improving AI agents can hallucinate non‑existent failures and create unnecessary guardrails, introducing a deterministic Counterfactual Fabrication…
cs.SE2026
When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems
Su Wang, Pin Qian, Yihang Chen +6
LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agentic AI systems: whether indivi…