Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Confidence Laundering in Agent Systems: Why Uncertainty Needs a Latent Carrier
Kaiwen Shi, Zheyuan Zhang, Han Bao +2
Modern agent systems can turn uncertainty into overconfidence. Fragile upstream decisions are often exposed to downstream components as clean intermediate artifacts, while the unce…
cs.AI2026
Drift-Bench: Diagnosing Cooperative Breakdowns in LLM Agents under Input Faults via Multi-Turn Interaction
Han Bao, Zheyuan Zhang, Pengcheng Jing +3
As Large Language Models transition to autonomous agents, user inputs frequently violate cooperative assumptions (e.g., implicit intent, missing parameters, false presuppositions,…