From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
Tianyu Chen, Chujia Hu, Wenjie Wang
The paper introduces Safety Sentry, a lightweight guard model for large language model agents that decides per action whether to execute, ask the user, or refuse, using a single de…
cs.AI2026
LPS-Bench: Benchmarking Safety Awareness of Computer-Use Agents in Long-Horizon Planning under Benign and Adversarial Scenarios
Tianyu Chen, Chujia Hu, Ge Gao +3
Computer-use agents (CUAs) that interact with real computer systems can perform automated tasks but face critical safety risks. Ambiguous instructions may trigger harmful actions,…