8 papers
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
Tom Sander, Hongyan Chang, Tomáš SouÄek +10
We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generation to restore output diversity,…
Learning to Watermark in the Latent Space of Generative Models
Sylvestre-Alvise Rebuffi, Tuan Tran, Valeriu Lacatusu +6
Existing approaches for watermarking AI-generated images often rely on post-hoc methods applied in pixel space, introducing computational overhead and potential visual artifacts. I…
How Good is Post-Hoc Watermarking With Language Model Rephrasing?
Pierre Fernandez, Tom Sander, Hady Elsahar +6
Generation-time text watermarking embeds statistical signals into text for traceability of AI-generated content. We explore *post-hoc watermarking* where an LLM rewrites existing t…
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
Redacted by arXiv
This document consolidates publicly reported technical details about Metas Llama 4 model family. It summarizes (i) released variants (Scout and Maverick) and the broader herd conte…
We Can Hide More Bits: The Unused Watermarking Capacity in Theory and in Practice
Aleksandar Petrov, Pierre Fernandez, Tomáš SouÄek +1
Despite rapid progress in deep learning-based image watermarking, the capacity of current robust methods remains limited to the scale of only a few hundred bits. Such plateauing pr…
Pixel Seal: Adversarial-only training for invisible image and video watermarking
Tomáš SouÄek, Pierre Fernandez, Hady Elsahar +5
Invisible watermarking is essential for tracing the provenance of digital content. However, training state-of-the-art models remains notoriously difficult, with current approaches…