Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Creative Collision: Directorial Persona Steering and Competition in Large Language Models
Subramanyam Sahoo, Justin Shenk
Activation steering has emerged as a powerful tool for shaping the behaviour of large language models at inference time, yet most prior work injects a \emph{single} semantic direct…
cs.CL2026
Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models
Subramanyam Sahoo
Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to questions that lie beyond their kno…
cs.CL2025
Catch Me If You Can: How Smaller Reasoning Models Pretend to Reason with Mathematical Fidelity
Subramanyam Sahoo, Vinija Jain, Saanidhya Vats +4
Current evaluation of mathematical reasoning in language models relies primarily on answer accuracy, potentially masking fundamental failures in logical computation. We introduce a…