4 papers
Leave No TRACE: Black-box Detection of Copyrighted Dataset Usage in Large Language Models via Watermarking
Jingqi Zhang, Ruibo Chen, Yingqing Yang +3
Large Language Models (LLMs) are increasingly fine-tuned on smaller, domain-specific datasets to improve downstream performance. These datasets often contain proprietary or copyrig…
Analyzing and Evaluating Unbiased Language Model Watermark
Yihan Wu, Xuehao Cui, Ruibo Chen +1
Verifying the authenticity of AI-generated text has become increasingly important with the rapid advancement of large language models, and unbiased watermarking has emerged as a pr…
An Ensemble Framework for Unbiased Language Model Watermarking
Yihan Wu, Ruibo Chen, Georgios Milis +1
As large language models become increasingly capable and widely deployed, verifying the provenance of machine-generated content is critical to ensuring trust, safety, and accountab…
A Watermark for Auto-Regressive Image Generation Models
Yihan Wu, Xuehao Cui, Ruibo Chen +2
The rapid evolution of image generation models has revolutionized visual content creation, enabling the synthesis of highly realistic and contextually accurate images for diverse a…