Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Practical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt Calibration
Wenjie Fu, Huandong Wang, Chen Gao +3
Membership Inference Attacks (MIA) aim to infer whether a target data record has been utilized for model training or not. Existing MIAs designed for large language models (LLMs) ca…
cs.CL2024
MIA-Tuner: Adapting Large Language Models as Pre-training Text Detector
Wenjie Fu, Huandong Wang, Chen Gao +3
The increasing parameters and expansive dataset of large language models (LLMs) highlight the urgent demand for a technical solution to audit the underlying privacy risks and copyr…