2 papers
cs.CR2026
Proof-of-Guardrail in AI Agents and What (Not) to Trust from It
Xisen Jin, Michael Duan, Qin Lin +4
As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which introduces a threat where safety measures…
cs.HC2026
Observable Social Life Spaces: Exploring User Interpretations of agent-side life context in human-agent interaction
Zihong He, Shuqin Wang, Songchen Zhou +4
Many AI agents are organized around instrumental "command-execution" interactions, where users primarily encounter agents through task requests and responses. Recent work on genera…