A Survey on the Explainability of Supervised Machine Learning
arXiv:2011.07876 · doi:10.1613/jair.1.12228
Abstract
Predictions obtained by, e.g., artificial neural networks have a high accuracy but humans often perceive the models as black boxes. Insights about the decision making are mostly opaque for humans. Particularly understanding the decision making in highly sensitive areas such as healthcare or fifinance, is of paramount importance. The decision-making behind the black boxes requires it to be more transparent, accountable, and understandable for humans. This survey paper provides essential definitions, an overview of the different principles and methodologies of explainable Supervised Machine Learning (SML). We conduct a state-of-the-art survey that reviews past and recent explainable SML approaches and classifies them according to the introduced definitions. Finally, we illustrate principles by means of an explanatory case study and discuss important future directions.
Accepted for publication at the Journal of Artificial Intelligence Research (JAIR)
References in corpus (31)
- Towards A Rigorous Science of Interpretable Machine Learning
- Methods for Interpreting and Understanding Deep Neural Networks
- Predictive learning via rule ensembles
- Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models
- A Survey on the Explainability of Supervised Machine Learning
- SmoothGrad: removing noise by adding noise
- Attention is not Explanation
- What Does Explainable AI Really Mean? A New Conceptualization of Perspectives
- Distilling a Neural Network Into a Soft Decision Tree
- Prototype selection for interpretable classification
- Metrics for Explainable AI: Challenges and Prospects
- NeuroRule: A Connectionist Approach to Data Mining
- Beyond Sparsity: Tree Regularization of Deep Models for Interpretability
- Gradients of Counterfactuals
- Quantifying Interpretability and Trust in Machine Learning Systems
- The Doctor Just Won't Accept That!
- Inverse Classification for Comparison-based Interpretability in Machine Learning
- Interpreting Classifiers through Attribute Interactions in Datasets
- Interpretable & Explorable Approximations of Black Box Models
- Methods and Models for Interpretable Linear Classification
- Explaining Trained Neural Networks with Semantic Web Technologies: First Steps
- Or's of And's for Interpretable Classification, with Application to Context-Aware Recommender Systems
- Explainable Restricted Boltzmann Machines for Collaborative Filtering
- Triaging moderate COVID-19 and other viral pneumonias from routine blood tests
- Neural Decision Trees
- Classification by Set Cover: The Prototype Vector Machine
- Towards Aggregating Weighted Feature Attributions
- An Investigation of COVID-19 Spreading Factors with Explainable AI Techniques
- Enhancing Transparency and Control when Drawing Data-Driven Inferences about Individuals
- Human-centric Transfer Learning Explanation via Knowledge Graph [Extended Abstract]
- REx: An Efficient Rule Generator