12 citations · 14 across the 3 of their papers we have counts for
3 papers
cs.LG2022★ 12 cited
Asymmetric Temperature Scaling Makes Larger Networks Teach Well Again
Xin-Chun Li, Wen-Shu Fan, Shaoming Song +4
Knowledge Distillation (KD) aims at transferring the knowledge of a well-performed neural network (the {\it teacher}) to a weaker one (the {\it student}). A peculiar phenomenon is…
cs.AI2022
On the Convergence Theory of Meta Reinforcement Learning with Personalized Policies
Haozhi Wang, Qing Wang, Yunfeng Shao +3
Modern meta-reinforcement learning (Meta-RL) methods are mainly developed based on model-agnostic meta-learning, which performs policy gradient steps across tasks to maximize polic…
cs.CV2022★ 2 cited
Federated Learning with Position-Aware Neurons
Xin-Chun Li, Yi-Chu Xu, Shaoming Song +4
Federated Learning (FL) fuses collaborative models from local nodes without centralizing users' data. The permutation invariance property of neural networks and the non-i.i.d. data…