4 papers
A Unified Game-Theoretic Interpretation of Adversarial Robustness
Jie Ren, Die Zhang, Yisen Wang +8
This paper provides a unified view to explain different adversarial attacks and defense methods, \emph{i.e.} the view of multi-order interactions between input variables of DNNs. B…
A Unified Game-Theoretic Interpretation of Adversarial Robustness
Jie Ren, Die Zhang, Yisen Wang +8
This paper provides a unified view to explain different adversarial attacks and defense methods, i.e. the view of multi-order interactions between input variables of DNNs. Based on…
Interpreting Multivariate Shapley Interactions in DNNs
Hao Zhang, Yichen Xie, Longjie Zheng +2
This paper aims to explain deep neural networks (DNNs) from the perspective of multivariate interactions. In this paper, we define and quantify the significance of interactions amo…
Building Interpretable Interaction Trees for Deep NLP Models
Die Zhang, Huilin Zhou, Hao Zhang +6
This paper proposes a method to disentangle and quantify interactions among words that are encoded inside a DNN for natural language processing. We construct a tree to encode salie…