From the 1 of 1 linked paper with an AI index.
1 paper
Jiajun Zhou, Zhaoxuan Ke, Jihang Ye +3
The paper presents AgentS4D, a sandboxed benchmark that evaluates runtime safety risks of large language model‑based workspace agents throughout their execution lifecycle, using a…