2 papers
cs.LG2026
Objective-Behavior Alignment: Diagnostics for MORL Policy Selection
Antonio Mone, Zuzanna Osika, Florian Felten +4
Real-world decision-making often requires optimizing multiple competing objectives simultaneously. In reinforcement learning (RL), this is typically addressed by combining reward s…
cs.LG2026
CoMI-IRL: Contrastive Multi-Intention Inverse Reinforcement Learning
Antonio Mone, Frans A. Oliehoek, Luciano Cavalcante Siebert
Inverse Reinforcement Learning (IRL) seeks to infer reward functions from expert demonstrations. When demonstrations originate from multiple experts with different intentions, the…