Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
CheckMIABench: Firm Foundations For Membership Inference Attacks on Language Models
Jeffrey G. Wang, Jason Wang, Marvin Li +1
Membership inference attacks (MIAs) are a canonical way to assess a machine learning model's privacy properties. Although several attempts have been made to evaluate MIAs on langua…
cs.LG2024
Bias Begets Bias: The Impact of Biased Embeddings on Diffusion Models
Sahil Kuchlous, Marvin Li, Jeffrey G. Wang
With the growing adoption of Text-to-Image (TTI) systems, the social biases of these models have come under increased scrutiny. Herein we conduct a systematic investigation of one…
cs.LG2023
MoPe: Model Perturbation-based Privacy Attacks on Language Models
Marvin Li, Jason Wang, Jeffrey Wang +1
Recent work has shown that Large Language Models (LLMs) can unintentionally leak sensitive information present in their training data. In this paper, we present Model Perturbations…