2 papers
cs.LG2026
Bayesian Anytime Pareto Set Identification for Multi-Objective Multi-Armed Bandits
Lennert Saerens, Bram Silue, Eleni Litsa +2
Identifying Pareto optimal solutions is critical to support multi-objective decision-making. We introduce the first anytime Multi-Objective Multi-Armed Bandit algorithm for the Par…
cs.LG2026
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
Bram Silue, Santiago Amaya-Corredor, Patrick Mannion +2
Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by inferring dense reward functions from…