1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jinghan Wang, Mengdi Wang, Lin F. Yang
This work considers the sample complexity of obtaining an ε-optimal policy in an average reward Markov Decision Process (AMDP), given access to a generative model (simu…