ProtoAttend: Attention-Based Prototypical Learning
arXiv:1902.06292
Abstract
We propose a novel inherently interpretable machine learning method that bases decisions on few relevant examples that we call prototypes. Our method, ProtoAttend, can be integrated into a wide range of neural network architectures including pre-trained models. It utilizes an attention mechanism that relates the encoded representations to samples in order to determine prototypes. The resulting model outperforms state of the art in three high impact problems without sacrificing accuracy of the original model: (1) it enables high-quality interpretability that outputs samples most relevant to the decision-making (i.e. a sample-based interpretability method); (2) it achieves state of the art confidence estimation by quantifying the mismatch across prototype labels; and (3) it obtains state of the art in distribution mismatch detection. All this can be achieved with minimal additional test time and a practically viable training time computational cost.
References in corpus (13)
- Prototypical Networks for Few-shot Learning
- What Uncertainties Do We Need in Bayesian Deep Learning for Computer Vision?
- On Calibration of Modern Neural Networks
- CatBoost: gradient boosting with categorical features support
- Understanding Black-box Predictions via Influence Functions
- This Looks Like That: Deep Learning for Interpretable Image Recognition
- Deep k-Nearest Neighbors: Towards Confident, Interpretable and Robust Deep Learning
- A Closer Look at Memorization in Deep Networks
- To Trust Or Not To Trust A Classifier
- Prototype selection for interpretable classification
- Deep Learning for Case-Based Reasoning through Prototypes: A Neural Network that Explains Its Predictions
- Incremental Few-Shot Learning with Attention Attractor Networks
- Bayesian Neural Networks