Showing cs.AIShow all
2 papers · 1 filter
cs.AI2019
Counterexample-Guided Strategy Improvement for POMDPs Using Recurrent Neural Networks
Steven Carr, Nils Jansen, Ralf Wimmer +3
We study strategy synthesis for partially observable Markov decision processes (POMDPs). The particular problem is to determine strategies that provably adhere to (probabilistic) t…
cs.AI2018
Human-in-the-Loop Synthesis for Partially Observable Markov Decision Processes
Steven Carr, Nils Jansen, Ralf Wimmer +2
We study planning problems where autonomous agents operate inside environments that are subject to uncertainties and not fully observable. Partially observable Markov decision proc…