1 paper · 1 filter
Hamish Flynn, Joe Watson, Ingmar Posner +1
We analyze the Bayesian regret of the Gaussian process posterior sampling reinforcement learning (GP-PSRL) algorithm. Posterior sampling is a heuristic for decision-making under un…