3 papers
cs.CL2025
Confidence is Not Competence
Debdeep Sanyal, Manya Pandey, Dhruv Kumar +2
Large language models (LLMs) often exhibit a puzzling disconnect between their asserted confidence and actual problem-solving competence. We offer a mechanistic account of this dec…
cs.CL2025
Policy Optimization Prefers The Path of Least Resistance
Debdeep Sanyal, Aakash Sen Sharma, Dhruv Kumar +2
Policy optimization (PO) algorithms are used to refine Large Language Models for complex, multi-step reasoning. Current state-of-the-art pipelines enforce a strict think-then-answe…
cs.LG2025
time2time: Causal Intervention in Hidden States to Simulate Rare Events in Time Series Foundation Models
Debdeep Sanyal, Aaryan Nagpal, Dhruv Kumar +2
While transformer-based foundation models excel at forecasting routine patterns, two questions remain: do they internalize semantic concepts such as market regimes, or merely fit c…