142 citations · 150 across the 2 of their papers we have counts for
1 paper · 1 filter
Daniel Hennes, Dustin Morrill, Shayegan Omidshafiei +8
Policy gradient and actor-critic algorithms form the basis of many commonly used training techniques in deep reinforcement learning. Using these algorithms in multiagent environmen…