2 papers
cs.RO2023
Get Back Here: Robust Imitation by Return-to-Distribution Planning
Geoffrey Cideron, Baruch Tabanpour, Sebastian Curi +6
We consider the Imitation Learning (IL) setup where expert data are not collected on the actual deployment environment but on a different version. To address the resulting distribu…
cs.LG2022
Safe Reinforcement Learning via Confidence-Based Filters
Sebastian Curi, Armin Lederer, Sandra Hirche +1
Ensuring safety is a crucial challenge when deploying reinforcement learning (RL) to real-world systems. We develop confidence-based safety filters, a control-theoretic approach fo…