1 paper
Kaito Ariu, Po-An Wang, Alexandre Proutiere +1
We study the policy testing problem in discounted Markov decision processes (MDPs) in the fixed-confidence setting under a generative model with static sampling. The goal is to dec…