Interpretability and Explainability: A Machine Learning Zoo Mini-tour
arXiv:2012.01805
Abstract
In this review, we examine the problem of designing interpretable and explainable machine learning models. Interpretability and explainability lie at the core of many machine learning and statistical applications in medicine, economics, law, and natural sciences. Although interpretability and explainability have escaped a clear universal definition, many techniques motivated by these properties have been developed over the recent 30 years with the focus currently shifting towards deep learning methods. In this review, we emphasise the divide between interpretability and explainability and illustrate these two different research directions with concrete examples of the state-of-the-art. The review is intended for a general machine learning audience with interest in exploring the problems of interpretation and explanation beyond logistic regression or random forest variable importance. This work is not an exhaustive literature survey, but rather a primer focusing selectively on certain lines of research which the authors found interesting or informative.
A preprint version of the 2023 WIREs Data Mining and Knowledge Discovery article
Cited by in corpus (9)
- Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
- Responsible and Regulatory Conform Machine Learning for Medicine: A Survey of Challenges and Solutions
- Discrete and continuous representations and processing in deep learning: Looking forward
- Symbolic Regression as Feature Engineering Method for Machine and Deep Learning Regression Tasks
- Can Requirements Engineering Support Explainable Artificial Intelligence? Towards a User-Centric Approach for Explainability Requirements
- Estimation of the masses in the Local Group by Gradient Boosted Decision Trees
- Is Machine Learning Unsafe and Irresponsible in Social Sciences? Paradoxes and Reconsidering from Recidivism Prediction Tasks
- Foreseeing the Impact of the Proposed AI Act on the Sustainability and Safety of Critical Infrastructures
- torchosr -- a PyTorch extension package for Open Set Recognition models evaluation in Python