244 citations · 991 across the 40 of their papers we have counts for
Showing 2020 · cs.LGShow all
2 papers · 2 filters
cs.LG2020
Reducing the Teacher-Student Gap via Spherical Knowledge Disitllation
Jia Guo, Minghao Chen, Yao Hu +3
Knowledge distillation aims at obtaining a compact and effective model by learning the mapping function from a much larger one. Due to the limited capacity of the student, the stud…
cs.LG2020
Do Wider Neural Networks Really Help Adversarial Robustness?
Boxi Wu, Jinghui Chen, Deng Cai +2
Adversarial training is a powerful type of defense against adversarial examples. Previous empirical results suggest that adversarial training requires wider networks for better per…