Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Every Token Leaves a Ripple in the Stream of Thought: Eliciting Model-Internal Token Saliency for Chain-of-Thought Compression
Tianyi Zhao, Yinhan He, Wendy Zheng +1
Chain-of-thought (CoT) reasoning improves multi-step problem solving, but long reasoning traces inflate inference cost. Token-level CoT compression reduces this cost by pruning ful…
cs.CL2026
Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs
Tianyi Zhao, Yinhan He, Wendy Zheng +2
Large language models are often not just wrong, but \emph{confidently wrong}: when they produce factually incorrect answers, they tend to verbalize overly high confidence rather th…
cs.CL2025
REHEARSE: Experiential Rehearsal for Verbal Confidence Calibration in Large Language Models
Ke Fang, Tianyi Zhao, Qianwen Wang +1
Large language models (LLMs) often express verbal confidence that is poorly aligned with actual correctness, limiting their reliability in safety-critical applications. Existing pr…