42 citations · 287 across the 71 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Putting on the Thinking Hats: A Survey on Chain of Thought Fine-tuning from the Perspective of Human Reasoning Mechanism
Xiaoshu Chen, Sihang Zhou, Ke Liang +5
Chain of thought (CoT) fine-tuning aims to endow large language models (LLMs) with reasoning capabilities by training them on curated reasoning traces. It leverages both supervised…
cs.CL2025
Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster
Xiao Chen, Sihang Zhou, Ke Liang +2
Chain-of-thought (CoT) distillation allows a large language model (LLM) to guide a small language model (SLM) in reasoning tasks. Existing methods train the SLM to learn the long r…
cs.CL2024
Distilling Reasoning Ability from Large Language Models with Adaptive Thinking
Xiaoshu Chen, Sihang Zhou, Ke Liang +1
Chain of thought finetuning (cot-finetuning) aims to endow small language models (SLM) with reasoning ability to improve their performance towards specific tasks by allowing them t…