1 paper · 1 filter
Xiao Chen, Sihang Zhou, Ke Liang +2
Chain-of-thought (CoT) distillation allows a large language model (LLM) to guide a small language model (SLM) in reasoning tasks. Existing methods train the SLM to learn the long r…