1 paper
Astrid Horn Brorholt, Maris F. L. Galesloot, Nils Jansen +2
Probabilistic shielding is a technique for safe reinforcement learning (RL). Typically, a static observer -- called the shield -- constrains the learning agent's actions to those f…