29 citations · 40 across the 9 of their papers we have counts for
4 papers
Use of LLMs for Illicit Purposes: Threats, Prevention Measures, and Vulnerabilities
Maximilian Mozes, Xuanli He, Bennett Kleinberg +1
Spurred by the recent rapid increase in the development and distribution of large language models (LLMs) across industry and academia, much recent work has drawn attention to safet…
IMBERT: Making BERT Immune to Insertion-based Backdoor Attacks
Xuanli He, Jun Wang, Benjamin Rubinstein +1
Backdoor attacks are an insidious security threat against machine learning models. Adversaries can manipulate the predictions of compromised models by inserting triggers into the t…
Koala: An Index for Quantifying Overlaps with Pre-training Corpora
Thuy-Trang Vu, Xuanli He, Gholamreza Haffari +1
In very recent years more attention has been placed on probing the role of pre-training data in Large Language Models (LLMs) downstream behaviour. Despite the importance, there is…
Protecting Intellectual Property of Language Generation APIs with Lexical Watermark
Xuanli He, Qiongkai Xu, Lingjuan Lyu +2
Nowadays, due to the breakthrough in natural language generation (NLG), including machine translation, document summarization, image captioning, etc NLG models have been encapsulat…