4 papers
The Alignment Flywheel: A Governance-Centric Hybrid MAS for Architecture-Agnostic Safety
Elias Malomgré, Pieter Simoens
Multi-agent systems provide mature methodologies for role decomposition, coordination, and normative governance, capabilities that remain essential as increasingly powerful autonom…
Interactionless Inverse Reinforcement Learning: A Data-Centric Framework for Durable Alignment
Elias Malomgré, Pieter Simoens
AI alignment is growing in importance, yet many current approaches learn safety behavior by directly modifying policy parameters, entangling normative constraints with the underlyi…
Ising Machines for Model Predictive Path Integral-Based Optimal Control
Lorin Werthen-Brabants, Pieter Simoens
We present a sampling-based Model Predictive Control (MPC) method that implements Model Predictive Path Integral (MPPI) as an \emph{Ising machine}, suitable for novel forms of prob…
Mixture of Autoencoder Experts Guidance using Unlabeled and Incomplete Data for Exploration in Reinforcement Learning
Elias Malomgré, Pieter Simoens
Recent trends in Reinforcement Learning (RL) highlight the need for agents to learn from reward-free interactions and alternative supervision signals, such as unlabeled or incomple…