4 citations · 6 across the 3 of their papers we have counts for
3 papers
math.OC2022
Algebraic optimization of sequential decision problems
Mareike Dressler, Marina Garrote-López, Guido Montúfar +2
We study the optimization of the expected long-term reward in finite partially observable Markov decision processes over the set of stationary stochastic policies. In the case of d…
cs.LG2022★ 2 cited
Solving infinite-horizon POMDPs with memoryless stochastic policies in state-action space
Johannes Müller, Guido Montúfar
Reward optimization in fully observable Markov decision processes is equivalent to a linear program over the polytope of state-action frequencies. Taking a similar perspective in t…
eess.SP2019★ 4 cited
Subjective Logic-based Identification of Markov Chains and Its Application to CAV's Safety
Johannes Müller, Thomas Griebel, Michael Gabb +1
A reliable estimation of the communication chan-nel which connects automated vehicles is an important steptowards the safety of connected and automated vehicles. The communication…