3 papers
cs.CR2026
Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content
Bing Liu, Shunping Wang, Yufan Zhu +5
This paper presents a survey and taxonomy of LLM fingerprinting and watermarking for identity, ownership verification, provenance, and generated-content attribution. Large language…
cs.CR2026
When Reasoning Leaks Membership: Membership Inference Attack on Black-box Large Reasoning Models
Ruihan Hu, Yu-Ming Shang, Wei Luo +2
Large Reasoning Models (LRMs) have rapidly gained prominence for their strong performance in solving complex tasks. Many modern black-box LRMs expose the intermediate reasoning tra…
cs.CL2025
Automated Detection of Pre-training Text in Black-box LLMs
Ruihan Hu, Yu-Ming Shang, Jiankun Peng +3
Detecting whether a given text is a member of the pre-training data of Large Language Models (LLMs) is crucial for ensuring data privacy and copyright protection. Most existing met…