2 papers
cs.LG2026
Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning
Omar Adalat, Edwin Hamel-De le Court, Francesco Belardinelli
Safe coordination problems surface in multi-agent reinforcement learning when global safety cannot be enforced by any agent unilaterally: the admissibility of one agent's action ma…
cs.AI2026
Robust Shielding for Safe Reinforcement Learning
Edwin Hamel-De le Court, Thom Badings, Alessandro Abate +2
Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, existing shielding techniques typ…