activity
20242026
collaborators

10 papers

stat.ML2026

Tightening the Score Matching Gap for Diffusion Models

Benjamin Dupuis, Tyler Farghly, Maxime Haddouche +2

Diffusion models (DMs) are a state-of-the-art generative method to approximately sample from an unknown distribution. Their training and evaluation primarily rely on an Evidence Lo…

stat.ML2026

Benign Overfitting Does Not Occur in Diffusion Models

Tyler Farghly, Benjamin Dupuis, Alain Durmus +1

Benign overfitting and double descent have come to shape our understanding of generalization in deep learning, establishing that overfitting is not only compatible with good genera…

stat.ML2026

Algorithm- and Data-Dependent Generalization Bounds for Diffusion Models

Benjamin Dupuis, Dario Shariatian, Maxime Haddouche +2

Score-based generative models (SGMs) have emerged as one of the most popular classes of generative models. A substantial body of work now exists on the analysis of SGMs, focusing e…

stat.ML2025

Differential privacy guarantees of Markov chain Monte Carlo algorithms

Andrea Bertazzi, Tim Johnston, Gareth O. Roberts +1

This paper aims to provide differential privacy (DP) guarantees for Markov chain Monte Carlo (MCMC) algorithms. In a first part, we establish DP guarantees on samples output by MCM…

cs.CV2025

Watermark Anything with Localized Messages

Tom Sander, Pierre Fernandez, Alain Durmus +2

Image watermarking methods are not tailored to handle small watermarked areas. This restricts applications in real-world scenarios where parts of the image may come from different…

cs.CR2025

Detecting Benchmark Contamination Through Watermarking

Tom Sander, Pierre Fernandez, Saeed Mahloujifar +2

Benchmark contamination poses a significant challenge to the reliability of Large Language Models (LLMs) evaluations, as it is difficult to assert whether a model has been trained…