Assessing high-order effects in feature importance via predictability decomposition
arXiv:2412.09964 · doi:10.1103/PhysRevE.111.L033301
Abstract
Leveraging the large body of work devoted in recent years to describe redundancy and synergy in multivariate interactions among random variables, we propose a novel approach to quantify cooperative effects in feature importance, one of the most used techniques for explainable artificial intelligence. In particular, we propose an adaptive version of a well-known metric of feature importance, named Leave One Covariate Out (LOCO), to disentangle high-order effects involving a given input feature in regression problems. LOCO is the reduction of the prediction error when the feature under consideration is added to the set of all the features used for regression. Instead of calculating the LOCO using all the features at hand, as in its standard version, our method searches for the multiplet of features that maximize LOCO and for the one that minimize it. This provides a decomposition of the LOCO as the sum of a two-body component and higher-order components (redundant and synergistic), also highlighting the features that contribute to building these high-order effects alongside the driving feature. We report the application to proton/pion discrimination from simulated detector measures by GEANT.
11 pages, 3 figures
References in corpus (7)
- Measuring Information Transfer
- The physics of higher-order interactions in complex systems
- Exploration of synergistic and redundant information sharing in static and dynamical Gaussian systems
- Visualizing the Feature Importance for Black Box Models
- Disentangling high-order mechanisms and high-order behaviours in complex systems
- Synergetic and redundant information flow detected by unnormalized Granger causality: application to resting state fMRI
- Redundant variables and Granger causality