11 papers
VIGIL: Runtime Enforcement of Behavioral Specifications in AI Agent Skills
Ying Li, Yanju Chen, Hongbo Wen +5
Agentic systems increasingly act through third-party skills, allowing model-generated decisions to affect files, communication channels, and cyber-physical devices. These skills of…
When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?
Yuan Tian, Bing Hu, Fang Wu +3
Think-with-image reasoning is emerging as a new inference paradigm for large vision-language models, but its safety implications remain poorly understood. Existing systems already…
Aligning Provenance with Authorization: A Dual-Graph Defense for LLM Agents
Peiran Wang, Ying Li, Yuan Tian
LLM-based agents are increasingly deployed in high-stakes scenarios such as email management, financial transactions, and code execution, where they interact with the external worl…
Reframing LLM Agent Security as an Agent-Human Interaction Problem
Peiran Wang, Ying Li, Yuan Tian
We argue that LLM agent security is fundamentally an agent-human interaction (AHI) problem, not a purely algorithmic one. To substantiate this position, we conduct a systematic ana…
APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents
Quan Zheng, Yan Gao, Shaobin He +6
LLM-agent simulation offers a flexible computational tool for studying population response trajectories that depend on scenario events, memory, demographics, and evolving social co…
No Attack Required: Semantic Fuzzing for Specification Violations in Agent Skills
Ying Li, Hongbo Wen, Yanju Chen +3
LLM-powered agents can silently delete documents, leak credentials, or transfer funds on a routine user request, not because the agent was attacked, but because the skill it invoke…