Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Multi-Aspect Knowledge Distillation for Language Model with Low-rank Factorization
Zihe Liu, Yulong Mao, Jinan Xu +2
Knowledge distillation is an effective technique for pre-trained language model compression. However, existing methods only focus on the knowledge distribution among layers, which…
cs.CL2024
Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation
Chen Liang, Zhifan Feng, Zihe Liu +4
Chain-of-thought prompting significantly boosts the reasoning ability of large language models but still faces three issues: hallucination problem, restricted interpretability, and…