Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Retry, Switch, or Abstain? Learning Strategy-Aware Tool-Use Policies via Controlled Error Injection
Chaoran Chen, Vy Nguyen, Ziji Zhang +7
Tool-using LLM agents are commonly trained and evaluated in environments where tool calls succeed reliably, yet deployed tools can fail transiently, persistently, or silently. Robu…
cs.AI2024
A Survey of Calibration Process for Black-Box LLMs
Liangru Xie, Hui Liu, Jingying Zeng +7
Large Language Models (LLMs) demonstrate remarkable performance in semantic understanding and generation, yet accurately assessing their output reliability remains a significant ch…