1 paper · 1 filter
Muhe Ding, Jianlong Wu, Xue Dong +4
Knowledge distillation is a mainstream algorithm in model compression by transferring knowledge from the larger model (teacher) to the smaller model (student) to improve the perfor…