3 papers
cs.CR2026
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
Tom Sander, Hongyan Chang, Tomáš SouÄek +10
We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generation to restore output diversity,…
cs.CR2025
Detecting Benchmark Contamination Through Watermarking
Tom Sander, Pierre Fernandez, Saeed Mahloujifar +2
Benchmark contamination poses a significant challenge to the reliability of Large Language Models (LLMs) evaluations, as it is difficult to assert whether a model has been trained…
cs.AI2024
Log-normal Mutations and their Use in Detecting Surreptitious Fake Images
Ismail Labiad, Thomas Bäck, Pierre Fernandez +5
In many cases, adversarial attacks are based on specialized algorithms specifically dedicated to attacking automatic image classifiers. These algorithms perform well, thanks to an…