2 papers
cs.CL2026
Freeze Deep, Train Shallow: Interpretable Layer Allocation for Continued Pre-Training
Yu-Hang Wu, Qin-Yuan Liu, Qiu-Yang Zhao +3
Selective layer-wise updates are essential for low-cost continued pre-training of Large Language Models (LLMs), yet determining which layers to freeze or train remains an empirical…
cs.CL2026
Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering
Li Zheng, Xin Zhang, Shuyi He +5
Accurate comprehension and controllable generation of emotion and rhetoric are pivotal for enhancing the reasoning capabilities of large language models (LLMs). Existing studies mo…