Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization
Yuhan Chen, Zhihua Tian, Mahavir Dabas +7
The performance of an LLM agent depends on the scaffold around a frozen model. A common way to improve that scaffold is to use a coding agent as an optimizer: it reads current scor…
cs.AI2026
Critic Experience Bank: Self-Evolving Step-Level Confidence Estimation for LLM Agents
Yaopei Zeng, Congchao Wang, JianHang Chen +3
LLM agents act in external environments where each action changes the state that later decisions condition on, and where a single wrong step can waste interaction budget or trigger…