Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Verify Smarter, Evolve Further: Efficient Harness Evolution through Behavior-Aware Verification
Jinghan Xu, Yikai Zhang, Aili Chen +3
Agent harnesses shape how language-model agents use instructions, tools, and runtime components, but adapting these harnesses requires costly verification. Existing propose-and-ver…
cs.AI2026
Beyond Single-Use Tokens: Durable Authorization State for Replay-Resistant LLM Agent Actions
Jinghan Xu, Longze Fan, Zeyuan Wang +2
Tool-using large language model agents frequently replan, retry failed operations, delegate tasks, and resume after crashes. These behaviors can cause one user authorization to be…