3 citations · 17 across the 14 of their papers we have counts for
1 paper · 2 filters
Hua Zheng, Wei Xie
Built on our previous study on green simulation assisted policy gradient (GS-PG) focusing on trajectory-based reuse, in this paper, we consider infinite-horizon Markov Decision Pro…