97 citations · 110 across the 7 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2019
Counterexample-Guided Strategy Improvement for POMDPs Using Recurrent Neural Networks
Steven Carr, Nils Jansen, Ralf Wimmer +3
We study strategy synthesis for partially observable Markov decision processes (POMDPs). The particular problem is to determine strategies that provably adhere to (probabilistic) t…
cs.AI2018
Safe Reinforcement Learning via Probabilistic Shields
Nils Jansen, Bettina Könighofer, Sebastian Junges +2
This paper targets the efficient construction of a safety shield for decision making in scenarios that incorporate uncertainty. Markov decision processes (MDPs) are prominent model…