answer conditioning 1chain of thought 1knowledge distillation 1large language models 1verifiable reasoning 1
From the 1 of 6 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models
Jungseob Lee, Seungyoon Lee, Suhyune Son +4
The paper shows that conditioning large language models on the correct answer when generating chains of thought harms the quality of distilled reasoning data, leading to large drop…
cs.CL2026
Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses
Sugyeong Eo, Heuiseok Lim
Although large language models (LLMs) have shown considerable progress in pragmatic language understanding, prior research has focused mainly on their comprehension of verbal behav…