Interpretable Machine Learning -- A Brief History, State-of-the-Art and Challenges
arXiv:2010.09337 · doi:10.1007/978-3-030-65965-3_28
Abstract
We present a brief history of the field of interpretable machine learning (IML), give an overview of state-of-the-art interpretation methods, and discuss challenges. Research in IML has boomed in recent years. As young as the field is, it has over 200 years old roots in regression modeling and rule-based machine learning, starting in the 1960s. Recently, many new IML methods have been proposed, many of them model-agnostic, but also interpretation techniques specific to deep learning and tree-based ensembles. IML methods either directly analyze model components, study sensitivity to input perturbations, or analyze local or global surrogate approximations of the ML model. The field approaches a state of readiness and stability, with many methods not only proposed in research, but also implemented in open-source software. But many important challenges remain for IML, such as dealing with dependent features, causal interpretation, and uncertainty estimation, which need to be resolved for its successful application to scientific problems. A further challenge is a missing rigorous definition of interpretability, which is accepted by the community. To address the challenges and advance the field, we urge to recall our roots of interpretable, data-driven modeling in statistics and (rule-based) ML, but also to consider other areas such as sensitivity analysis, causal inference, and the social sciences.
References in corpus (13)
- Deep Learning in Neural Networks: An Overview
- Towards A Rigorous Science of Interpretable Machine Learning
- Predictive learning via rule ensembles
- Towards Explainable Artificial Intelligence
- Variable importance in binary regression trees and forests
- Distilling a Neural Network Into a Soft Decision Tree
- Metrics for Explainable AI: Challenges and Prospects
- Feature relevance quantification in explainable AI: A causal problem
- Interpretable Predictions of Tree-based Ensembles via Actionable Feature Tweaking
- Relative Feature Importance
- A study of data and label shift in the LIME framework
- audioLIME: Listenable Explanations Using Source Separation
- MeLIME: Meaningful Local Explanation for Machine Learning Models
Cited by in corpus (12)
- Choice modelling in the age of machine learning -- discussion paper
- A Prescriptive Learning Analytics Framework: Beyond Predictive Modelling and onto Explainable AI with Prescriptive Analytics and ChatGPT
- Verifying Controllers with Convolutional Neural Network-based Perception: A Case for Intelligible, Safe, and Precise Abstractions
- A Survey on Neural Network Interpretability
- Scientific Inference With Interpretable Machine Learning: Analyzing Models to Learn About Real-World Phenomena
- Transforming Feature Space to Interpret Machine Learning Models
- Explainable Ensemble-Based Machine Learning Models for Detecting the Presence of Cirrhosis in Hepatitis C Patients
- Explaining classifiers to understand coarse-grained models
- Explanations as Programs in Probabilistic Logic Programming
- Interpretable machine-learning identification of the crossover from subradiance to superradiance in an atomic array
- Neural-to-Tree Policy Distillation with Policy Improvement Criterion
- BDD4BNN: A BDD-based Quantitative Analysis Framework for Binarized Neural Networks