3 papers
cs.AI2025
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
Vardhan Palod, Karthik Valmeekam, Kaya Stechly +1
Intermediate token generation (ITG), where a model produces output before the solution, has been proposed as a method to improve the performance of language models on reasoning tas…
cs.LG2025
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
Soumya Rani Samineni, Durgesh Kalwar, Karthik Valmeekam +2
Reinforcement learning based post-training of large language models (LLMs) has recently gained attention, particularly following the release of DeepSeek R1, which applied GRPO for…
cs.CL2024
Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach
Atharva Gundawar, Karthik Valmeekam, Mudit Verma +1
Previous work has attempted to boost Large Language Model (LLM) performance on planning and scheduling tasks through a variety of prompt engineering techniques. While these methods…