2 papers
cs.LG2025
Viability of Future Actions: Robust Safety in Reinforcement Learning via Entropy Regularization
Pierre-François Massiani, Alexander von Rohr, Lukas Haverbeck +1
Despite the many recent advances in reinforcement learning (RL), the question of learning policies that robustly satisfy state constraints under unknown disturbances remains open.…
cs.LG2024
On the Consistency of Kernel Methods with Dependent Observations
Pierre-François Massiani, Sebastian Trimpe, Friedrich Solowjow
The consistency of a learning method is usually established under the assumption that the observations are a realization of an independent and identically distributed (i.i.d.) or m…