11 citations · 12 across the 6 of their papers we have counts for
1 paper · 1 filter
Shuang Qiu, Xiaohan Wei, Jieping Ye +2
While single-agent policy optimization in a fixed environment has attracted a lot of research attention recently in the reinforcement learning community, much less is known theoret…