2 papers
cs.LG2026
Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning
Omar Adalat, Edwin Hamel-De le Court, Francesco Belardinelli
Safe coordination problems surface in multi-agent reinforcement learning when global safety cannot be enforced by any agent unilaterally: the admissibility of one agent's action ma…
cs.LG2025
Expressive Temporal Specifications for Reward Monitoring
Omar Adalat, Francesco Belardinelli
Specifying informative and dense reward functions remains a pivotal challenge in Reinforcement Learning, as it directly affects the efficiency of agent training. In this work, we h…