2 papers
cs.LG2025
Kernel-Smoothed Scores for Denoising Diffusion: A Bias-Variance Study
Franck Gabriel, François Ged, Maria Han Veiga +1
Diffusion models now set the benchmark in high-fidelity generative sampling, yet they can, in principle, be prone to memorization. In this case, their learned score overfits the fi…
cs.LG2024
Matryoshka Policy Gradient for Entropy-Regularized RL: Convergence and Global Optimality
François Ged, Maria Han Veiga
A novel Policy Gradient (PG) algorithm, called (MPG), is introduced and studied, in the context of fixed-horizon max-entropy reinforcement lea…