Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Robust Shielding for Safe Reinforcement Learning
Edwin Hamel-De le Court, Thom Badings, Alessandro Abate +2
Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, existing shielding techniques typ…
cs.AI2025
Best-Effort Policies for Robust Markov Decision Processes
Alessandro Abate, Thom Badings, Giuseppe De Giacomo +1
We study the common generalization of Markov decision processes (MDPs) with sets of transition probabilities, known as robust MDPs (RMDPs). A standard goal in RMDPs is to compute a…