Qualitative Analysis of Partially-observable Markov Decision Processes
arXiv:0909.1645 · doi:10.1007/978-3-642-15155-2_24
Abstract
We study observation-based strategies for partially-observable Markov decision processes (POMDPs) with omega-regular objectives. An observation-based strategy relies on partial information about the history of a play, namely, on the past sequence of observations. We consider the qualitative analysis problem: given a POMDP with an omega-regular objective, whether there is an observation-based strategy to achieve the objective with probability~1 (almost-sure winning), or with positive probability (positive winning). Our main results are twofold. First, we present a complete picture of the computational complexity of the qualitative analysis of POMDP s with parity objectives (a canonical form to express omega-regular objectives) and its subclasses. Our contribution consists in establishing several upper and lower bounds that were not known in literature. Second, we present optimal bounds (matching upper and lower bounds) on the memory required by pure and randomized observation-based strategies for the qualitative analysis of POMDP s with parity objectives and its subclasses.
References in corpus (3)
Cited by in corpus (10)
- Qualitative Analysis of Partially-observable Markov Decision Processes
- Stochastic Finite State Control of POMDPs with LTL Specifications
- Games on Graphs: From Logic and Automata to Algorithms
- Reachability Analysis of Quantum Markov Decision Processes
- Probabilistic Weighted Automata
- Probabilistic Bisimulation: Naturally on Distributions
- Memory Lens: How Much Memory Does an Agent Use?
- Stochastic Shortest Path with Energy Constraints in POMDPs
- Equivalence of Games with Probabilistic Uncertainty and Partial-observation Games
- Probabilistic Opacity for Markov Decision Processes