From the 1 of 4 linked papers with an AI index.
4 papers
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
Tianyu Chen, Chujia Hu, Wenjie Wang
The paper introduces Safety Sentry, a lightweight guard model for large language model agents that decides per action whether to execute, ask the user, or refuse, using a single de…
Safeguarding Multimodal Knowledge Copyright in the RAG-as-a-Service Environment
Tianyu Chen, Jian Lou, Wenjie Wang
As Retrieval-Augmented Generation (RAG) evolves into service-oriented platforms (Rag-as-a-Service) with shared knowledge bases, protecting the copyright of contributed data becomes…
A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)
Tianyu Chen, Dongrui Liu, Xia Hu +2
Clawdbot is a self-hosted, tool-using personal AI agent with a broad action space spanning local execution and web-mediated workflows, which raises heightened safety and security c…
LPS-Bench: Benchmarking Safety Awareness of Computer-Use Agents in Long-Horizon Planning under Benign and Adversarial Scenarios
Tianyu Chen, Chujia Hu, Ge Gao +3
Computer-use agents (CUAs) that interact with real computer systems can perform automated tasks but face critical safety risks. Ambiguous instructions may trigger harmful actions,…