3 papers
cs.CL2026
Failing to Falsify: Evaluating and Mitigating Confirmation Bias in Language Models
Ayush Rajesh Jhaveri, Anthony GX-Chen, Ilia Sucholutsky +1
Confirmation bias, the tendency to seek evidence that supports rather than challenges one's belief, hinders one's reasoning ability. We examine whether large language models (LLMs)…
cs.CL2025
Interpreting and Mitigating Unwanted Uncertainty in LLMs
Tiasa Singha Roy, Ayush Rajesh Jhaveri, Ilias Triantafyllopoulos
Despite their impressive capabilities, Large Language Models (LLMs) exhibit unwanted uncertainty, a phenomenon where a model changes a previously correct answer into an incorrect o…
cs.CL2025
Can LLMs Math? -- Exploring the Pitfalls in Mathematical Reasoning
Tiasa Singha Roy, Aditeya Baral, Ayush Rajesh Jhaveri +1
Large language models (LLMs) demonstrate considerable potential in various natural language tasks but face significant challenges in mathematical reasoning, particularly in executi…