Showing 2024Show all
2 papers · 1 filter
stat.ML2024
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
Daniel Adelman, Cagla Keceli, Alba V. Olivares-Nadal
This paper develops a viable notion of learning for sampling-based algorithms that applies in broader settings than previously considered. More specifically, we model a discounted…
math.OC2024
Measurized Markov Decision Processes
Daniel Adelman, Alba V. Olivares-Nadal
In this paper, we explore lifting Markov Decision Processes (MDPs) to the space of probability measures and consider the so-called measurized MDPs: deterministic processes where st…