2 papers
cs.IT2026
Covert Multi-bit LLM Watermarking: An Information Theory and Coding Approach
Sidong Guo, Tyler Kann, Teodora Baluta +1
We study the problem of multi-bit watermarking for large language models (LLMs). We introduce a block-autoregressive model inspired by multi-token prediction, in which the encoder…
cs.CR2025
Model Provenance Testing for Large Language Models
Ivica Nikolic, Teodora Baluta, Prateek Saxena
Large language models are increasingly customized through fine-tuning and other adaptations, creating challenges in enforcing licensing terms and managing downstream impacts. Track…