3 papers
math.OC2022
A Non-asymptotic Analysis of Non-parametric Temporal-Difference Learning
Eloïse Berthier, Ziad Kobeissi, Francis Bach
Temporal-difference learning is a popular algorithm for policy evaluation. In this paper, we study the convergence of the regularized non-parametric TD(0) algorithm, in both the in…
math.OC2021
Infinite-Dimensional Sums-of-Squares for Optimal Control
Eloïse Berthier, Justin Carpentier, Alessandro Rudi +1
We introduce an approximation method to solve an optimal control problem via the Lagrange dual of its weak formulation. It is based on a sum-of-squares representation of the Hamilt…
cs.LG2019
Amplifying Rényi Differential Privacy via Shuffling
Eloïse Berthier, Sai Praneeth Karimireddy
Differential privacy is a useful tool to build machine learning models which do not release too much information about the training data. We study the Rényi differential privacy of…