Explainable Matrix -- Visualization for Global and Local Interpretability of Random Forest Classification Ensembles
arXiv:2005.04289 · doi:10.1109/TVCG.2020.3030354
Abstract
Over the past decades, classification models have proven to be essential machine learning tools given their potential and applicability in various domains. In these years, the north of the majority of the researchers had been to improve quantitative metrics, notwithstanding the lack of information about models' decisions such metrics convey. This paradigm has recently shifted, and strategies beyond tables and numbers to assist in interpreting models' decisions are increasing in importance. Part of this trend, visualization techniques have been extensively used to support classification models' interpretability, with a significant focus on rule-based models. Despite the advances, the existing approaches present limitations in terms of visual scalability, and the visualization of large and complex models, such as the ones produced by the Random Forest (RF) technique, remains a challenge. In this paper, we propose Explainable Matrix (ExMatrix), a novel visualization method for RF interpretability that can handle models with massive quantities of rules. It employs a simple yet powerful matrix-like visual metaphor, where rows are rules, columns are features, and cells are rules predicates, enabling the analysis of entire models and auditing classification results. ExMatrix applicability is confirmed via different examples, showing how it can be used in practice to promote RF models interpretability.
IEEE VIS VAST 2020
References in corpus (4)
Cited by in corpus (13)
- Diagnosing Ensemble Few-Shot Classifiers
- VisEvol: Visual Analytics to Support Hyperparameter Search through Evolutionary Optimization
- TimberTrek: Exploring and Curating Sparse Decision Trees with Interactive Visualization
- Explainable automatic industrial carbon footprint estimation from bank transaction classification using natural language processing
- VisRuler: Visual Analytics for Extracting Decision Rules from Bagged and Boosted Decision Trees
- The Pattern is in the Details: An Evaluation of Interaction Techniques for Locating, Searching, and Contextualizing Details in Multivariate Matrix Visualizations
- Local Multi-Label Explanations for Random Forest
- Automatic explanation of the classification of Spanish legal judgments in jurisdiction-dependent law categories with tree estimators
- Decision Predicate Graphs: Enhancing Interpretability in Tree Ensembles
- DeforestVis: Behavior Analysis of Machine Learning Models with Surrogate Decision Stumps
- From global to local MDI variable importances for random forests and when they are Shapley values
- Visualizing Rule Sets: Exploration and Validation of a Design Space
- Facilitating Machine Learning Model Comparison and Explanation Through A Radial Visualisation