12 citations · 14 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 12 cited
Asymmetric Temperature Scaling Makes Larger Networks Teach Well Again
Xin-Chun Li, Wen-Shu Fan, Shaoming Song +4
Knowledge Distillation (KD) aims at transferring the knowledge of a well-performed neural network (the {\it teacher}) to a weaker one (the {\it student}). A peculiar phenomenon is…
cs.CV2022★ 2 cited
Federated Learning with Position-Aware Neurons
Xin-Chun Li, Yi-Chu Xu, Shaoming Song +4
Federated Learning (FL) fuses collaborative models from local nodes without centralizing users' data. The permutation invariance property of neural networks and the non-i.i.d. data…