2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CL2025
Delta Knowledge Distillation for Large Language Models
Yihan Cao, Yanbin Kang, Zhengming Xing +1
Knowledge distillation (KD) is a widely adopted approach for compressing large neural networks by transferring knowledge from a large teacher model to a smaller student model. In t…
cs.CL2023★ 2 cited
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
Yihan Cao, Yanbin Kang, Chi Wang +1
Large language models (LLMs) are initially pretrained for broad capabilities and then finetuned with instruction-following datasets to improve their performance in interacting with…