paper

Membership Inference Attacks and Generalization: A Causal Perspective

arXiv:2209.08615 · doi:10.1145/3548606.3560694

Abstract

Membership inference (MI) attacks highlight a privacy weakness in present stochastic training methods for neural networks. It is not well understood, however, why they arise. Are they a natural consequence of imperfect generalization only? Which underlying causes should we address during training to mitigate these attacks? Towards answering such questions, we propose the first approach to explain MI attacks and their connection to generalization based on principled causal reasoning. We offer causal graphs that quantitatively explain the observed MI attack performance achieved for attack variants. We refute several prior non-quantitative hypotheses that over-simplify or over-estimate the influence of underlying causes, thereby failing to capture the complex interplay between several factors. Our causal models also show a new connection between generalization and MI attacks via their shared causal factors. Our causal models have high predictive power (), i.e., their analytical predictions match with observations in unseen experiments often, which makes analysis via them a pragmatic alternative.

26 pages, 15 figures; added CC-license block icons and links, typos corrected, added reference to Github

References in corpus (4)

Membership Inference Attacks and Generalization: A Causal Perspective · wovepaper