1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Yudan Wang, Yue Wang, Yi Zhou +1
Actor-critic (AC) is a powerful method for learning an optimal policy in reinforcement learning, where the critic uses algorithms, e.g., temporal difference (TD) learning with func…