2 papers
cs.CL2026
Calibration Drift Under Reasoning: How Chain-of-Thought Budgets Induce Overconfidence in Large Language Models
Prakul Sunil Hiremath, Harshit R. Hiremath
The ability of large language models (LLMs) to express calibrated uncertainty is important for safe deployment. Chain-of-thought (CoT) reasoning is widely used to improve accuracy…
cs.CR2026
Sensitivity Uncertainty Alignment in Large Language Models
Prakul Sunil Hiremath, Harshit R. Hiremath
We propose Sensitivity-Uncertainty Alignment (SUA), a framework for analyzing failures of large language models under adversarial and ambiguous inputs. We argue that adversarial se…