4 citations · 4 across the 1 of their papers we have counts for
1 paper · 1 filter
Yaqi Duan, Mengdi Wang, Martin J. Wainwright
We study methods based on reproducing kernel Hilbert spaces for estimating the value function of an infinite-horizon discounted Markov reward process (MRP). We study a regularized…