3 papers
cs.AI2026
Masked Distillation: Internalizing the Chain-of-Thought in Language Models
Durgesh Kalwar, Vardhan Palod, Subbarao Kambhampati
Large Reasoning Models (LRMs) produce long, explicit chains of intermediate steps before generating a final answer at inference time. These intermediate traces dominate latency, me…
cs.AI2025
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
Vardhan Palod, Karthik Valmeekam, Kaya Stechly +1
Intermediate token generation (ITG), where a model produces output before the solution, has been proposed as a method to improve the performance of language models on reasoning tas…
q-bio.NC2025
Discounting and Drug Seeking in Biological Hierarchical Reinforcement Learning
Vardhan Palod, Pranav Mahajan, Veeky Baths +1
Despite a strong desire to quit, individuals with long-term substance use disorder (SUD) often struggle to resist drug use, even when aware of its harmful consequences. This discon…