4 citations · 4 across the 1 of their papers we have counts for
1 paper
Wenlong Mou, Zheng Wen, Xi Chen
We study the optimal sample complexity in large-scale Reinforcement Learning (RL) problems with policy space generalization, i.e. the agent has a prior knowledge that the optimal p…