1 citations · 1 across the 11 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
DualStake: Dual-Path Confidence Calibration in Deep Research Agents
Yinuo Xu, Yuwei Liang, Jianjie Cheng +4
Deep Research agents tackle knowledge-intensive tasks through multi-round retrieval and decision-oriented generation. However, these agents suffer from severe overconfidence, makin…
cs.CL2025
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models
Yanbo Wang, Yongcan Yu, Jian Liang +1
The development of Long-CoT reasoning has advanced LLM performance across various tasks, including language understanding, complex problem solving, and code generation. This paradi…