2 papers
cs.CL2025
Automated Detection of Pre-training Text in Black-box LLMs
Ruihan Hu, Yu-Ming Shang, Jiankun Peng +3
Detecting whether a given text is a member of the pre-training data of Large Language Models (LLMs) is crucial for ensuring data privacy and copyright protection. Most existing met…
cs.SI2024
Tracing the Unseen: Uncovering Human Trafficking Patterns in Job Listings
Siyi Zhou, Jiankun Peng, Emilio Ferrara
In the shadow of the digital revolution, the insidious issue of human trafficking has found new breeding grounds within the realms of social media and online job boards. Previous r…