1 paper
Shangziqi Zhao, Jiahao Yuan, Jinyang Wu +3
Long chain-of-thought (Long-CoT) reasoning improves accuracy in LLMs, yet its verbose, self-reflective style often hinders effective distillation into small language models (SLMs).…