1 paper
Levin Hornischer, Hannes Leitgeb
We propose a new interpretability method for neural networks, which is based on a novel mathematico-philosophical theory of reasons. Our method computes a vector for each neuron, c…