3 papers
cs.LG2026
SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer
Nathan Samuel de Lara, Florian Shkurti
Modern offline Reinforcement Learning (RL) methods find performant actor-critics, however, fine-tuning these actor-critics online with value-based RL algorithms typically causes im…
cs.RO2025
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
Hossein Goli, Michael Gimelfarb, Nathan Samuel de Lara +3
Off-policy evaluation (OPE) estimates the performance of a target policy using offline data collected from a behavior policy, and is crucial in domains such as robotics or healthca…
cs.LG2023
A Consistent Diffusion-Based Algorithm for Semi-Supervised Graph Learning
Thomas Bonald, Nathan de Lara
The task of semi-supervised classification aims at assigning labels to all nodes of a graph based on the labels known for a few nodes, called the seeds. One of the most popular alg…