2 papers
cs.SE2026
Keep Evaluation Fair: Detecting Data Leakage in Code Generation Benchmarks via Membership Inference Attacks
Dongdong Zhao, Jian Chen, Guancheng Lin +3
Code generation benchmarks are widely used to evaluate Large Language Models (LLMs), but benchmark data leakage into training sets can inflate performance and undermine evaluation…
cs.CR2026
From Classification to Consistent Templates: Multiple Permuted-Label Classifier Encoding for Biometric Template Protection
Baogang Song, Zhongshu Zhao, Qianrong Zheng +2
Biometric template protection (BTP) must secure stored templates while tolerating intra-class variations. Existing methods rely on protected-domain similarity matching, error corre…