4 papers
Provable Joint Decontamination for Benchmarking Multiple Large Language Models
Zhenlong Liu, Hao Zeng, Hongxin Wei
Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audited models, reported performa…
Provable Training Data Identification for Large Language Models
Zhenlong Liu, Hao Zeng, Weiran Huang +1
Identifying training data of large-scale models is critical for copyright litigation, privacy auditing, and ensuring fair evaluation. However, existing works typically treat this t…
How does Bayesian Sampling help Membership Inference Attacks?
Zhenlong Liu, Wenyu Jiang, Feng Zhou +1
Membership Inference Attacks (MIAs) aim to estimate whether a specific data point was used in the training of a given model. Existing state-of-the-art attacks typically rely on tra…
A Modulo Sampling Hardware Prototype and Reconstruction Algorithm Evaluation
Jiang Zhu, Junnan Ma, Zhenlong Liu +3
Analog-to-digital converters (ADCs) play a vital important role in any devices via manipulating analog signals in a digital manner. Given that the amplitude of the signal exceeds t…