Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention
Wenye Lin, Kai Han
Injecting new reasoning knowledge into Large Language Models (LLMs) via post-training often induces catastrophic forgetting. Recent studies emphasize the importance of on-policy da…
cs.CL2026
VersatileFFN: Achieving Parameter Efficiency in LLMs via Adaptive Wide-and-Deep Reuse
Ying Nie, Kai Han, Hongguang Li +5
The rapid scaling of Large Language Models (LLMs) has achieved remarkable performance, but it also leads to prohibitive memory costs. Existing parameter-efficient approaches such a…