calibration 1experience replay 1LLM agents 1self-evolving critic 1step-level confidence 1training-free methods 1
From the 1 of 7 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Critic Experience Bank: Self-Evolving Step-Level Confidence Estimation for LLM Agents
Yaopei Zeng, Congchao Wang, JianHang Chen +3
The paper proposes the Critic Experience Bank, a training-free framework that lets large language model agents estimate confidence for each action by storing and retrieving past st…
cs.AI2026
Forced Deferral: Manipulating Routing Decisions in Multimodal LLM Cascades
Zhongye Liu, Yaopei Zeng, Yurui Chang +1
While multimodal large language models (MLLMs) have shown strong visual reasoning abilities, serving a large model for every query is computationally expensive. MLLM cascades mitig…