4 citations · 9 across the 5 of their papers we have counts for
1 paper · 1 filter
Ying Wen, Hui Chen, Yaodong Yang +4
Trust region methods are widely applied in single-agent reinforcement learning problems due to their monotonic performance-improvement guarantee at every iteration. Nonetheless, wh…