Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
MI-Distillation: Selecting from Model-Interpolated Instruct-Reasoning Data Spectrum for Chain-of-Thought Distillation
Yangsong Lan, Renkai Hu, HongKai Zheng +4
Recent advances in large reasoning models (LRMs) have shown strong performance on complex problems through long chain-of-thought (Long CoT) reasoning. However, distilling such traj…
cs.CL2026
CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning
Yangsong Lan, Hongliang Dai, Piji Li
Long Chain-of-Thought (CoT) reasoning is pivotal for the success of recent reasoning models but suffers from high computational overhead and latency. While prior works attempt to c…
cs.CL2024
5W1H Extraction With Large Language Models
Yang Cao, Yangsong Lan, Feiyan Zhai +1
The extraction of essential news elements through the 5W1H framework (\textit{What}, \textit{When}, \textit{Where}, \textit{Why}, \textit{Who}, and \textit{How}) is critical for ev…