activity
20242026
most citedA Transformer Model for Predicting Chemical Products from Generic SMARTS Templates with Data Augmentation

1 citations · 2 across the 10 of their papers we have counts for

collaborators
Showing cs.LGShow all

8 papers · 1 filter

cs.LG2026

Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment

Mathieu Petitbois, Rémy Portelas, Sylvain Lamprier

We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. In this setting, aligning style with hi…

cs.LG2025

HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents

Thomas Carta, Clément Romac, Loris Gaven +3

We study goal-conditioned reinforcement learning in partially observable environments with sparse rewards and large, structured goal spaces. In such settings, complex goals often r…

cs.LG2025

Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning

Nicolas Castanet, Olivier Sigaud, Sylvain Lamprier

Goal-Conditioned Reinforcement Learning (GCRL) enables agents to autonomously acquire diverse behaviors, but faces major challenges in visual environments due to high-dimensional,…

cs.LG2025

Offline Learning of Controllable Diverse Behaviors

Mathieu Petitbois, Rémy Portelas, Sylvain Lamprier +1

Imitation Learning (IL) techniques aim to replicate human behaviors in specific tasks. While IL has gained prominence due to its effectiveness and efficiency, traditional methods o…

cs.LG2025

A Transformer Model for Predicting Chemical Products from Generic SMARTS Templates with Data Augmentation

Derin Ozer, Sylvain Lamprier, Thomas Cauchy +2

The accurate prediction of chemical reaction outcomes is a major challenge in computational chemistry. Current models rely heavily on either highly specific reaction templates or t…

cs.LG2024

Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning

Alexi Canesse, Mathieu Petitbois, Ludovic Denoyer +2

Offline Reinforcement Learning (RL) has emerged as a powerful alternative to imitation learning for behavior modeling in various domains, particularly in complex navigation tasks.…