2 papers
cs.LG2020
Neighbourhood Distillation: On the benefits of non end-to-end distillation
Laëtitia Shao, Max Moroz, Elad Eban +1
End-to-end training with back propagation is the standard method for training deep neural networks. However, as networks become deeper and bigger, end-to-end training becomes more…
cs.LG2020
Understanding Classifier Mistakes with Generative Models
Laëtitia Shao, Yang Song, Stefano Ermon
Although deep neural networks are effective on supervised learning tasks, they have been shown to be brittle. They are prone to overfitting on their training distribution and are e…