2 papers
cs.LG2026
Learning to maintain safety through expert demonstrations in settings with unknown constraints: A Q-learning perspective
George Papadopoulos, George A. Vouros
Given a set of trajectories demonstrating the execution of a task safely in a constrained MDP with observable rewards but with unknown constraints and non-observable costs, we aim…
cs.LG2025
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
George Papadopoulos, George A. Vouros
This article introduces an imitation learning method for learning maximum entropy policies that comply with constraints demonstrated by expert trajectories executing a task. The fo…