127 citations · 149 across the 4 of their papers we have counts for
4 papers
Fast Value Iteration for Goal-Directed Markov Decision Processes
Nevin Lianwen Zhang, Weihong Zhang
Planning problems where effects of actions are non-deterministic can be modeled as Markov decision processes. Planning problems are usually goal-directed. This paper proposes sever…
A Method for Speeding Up Value Iteration in Partially Observable Markov Decision Processes
Nevin Lianwen Zhang, Stephen S. Lee, Weihong Zhang
We present a technique for speeding up the convergence of value iteration for partially observable Markov decisions processes (POMDPs). The underlying idea is similar to that behin…
Restricted Value Iteration: Theory and Algorithms
N. L. Zhang, W. Zhang
Value iteration is a popular algorithm for finding near optimal policies for POMDPs. It is inefficient due to the need to account for the entire belief space, which necessitates th…
Speeding Up the Convergence of Value Iteration in Partially Observable Markov Decision Processes
N. L. Zhang, W. Zhang
Partially observable Markov decision processes (POMDPs) have recently become popular among many AI researchers because they serve as a natural model for planning under uncertainty.…