Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Bayesian learning for the stochastic shortest path problem
Chon Wai Ho, Sumeetpal S. Singh, Jiaqi Guo
Sequential decision-making problems are often modelled as a Markov decision process (MDP). We focus on the stochastic shortest path (SSP) problem, which is an infinite-horizon undi…
stat.ML2025
Bayesian learning of the optimal action-value function in a Markov decision process
Jiaqi Guo, Chon Wai Ho, Sumeetpal S. Singh
The Markov Decision Process (MDP) is a popular framework for sequential decision-making problems, and uncertainty quantification is an essential component of it to learn optimal de…