paper

Sparse Point-wise Privacy Leakage: Mechanism Design and Fundamental Limits

arXiv:2601.07523

Abstract

We study an information-theoretic privacy mechanism design problem, where an agent observes useful data that is arbitrarily correlated with sensitive data , and design disclosed data generated from (the agent has no direct access to ). We introduce \emph{sparse point-wise privacy leakage}, a worst-case privacy criterion that enforces two simultaneous constraints for every disclosed symbol : (i) may be correlated with at most realizations of , and (ii) the total leakage toward those realizations is bounded. In the high-privacy regime, we use concepts from information geometry to obtain a local quadratic approximation of mutual information which measures utility between and . When the leakage matrix is invertible, this approximation reduces the design problem to a sparse quadratic maximization, known as the Rayleigh-quotient problem, with an constraint. We further show that, for the approximated problem, one can without loss of optimality restrict attention to a binary released variable with a uniform distribution. For small alphabet sizes, the exact sparsity-constrained optimum can be computed via combinatorial support enumeration, which quickly becomes intractable as the dimension grows. For general dimensions, the resulting sparse Rayleigh-quotient maximization is NP-hard and closely related to sparse principal component analysis (PCA). We propose a convex semidefinite programming (SDP) relaxation that is solvable in polynomial time and provides a tractable surrogate for the NP-hard design, together with a simple rounding procedure to recover a feasible leakage direction. We also identify a sparsity threshold beyond which the sparse optimum saturates at the unconstrained spectral value and the SDP relaxation becomes tight.

Sparse Point-wise Privacy Leakage: Mechanism Design and Fundamental Limits · wovepaper