activity
20162021
most citedPolicy Gradient Methods for Reinforcement Learning with Function Approximation and Action-Dependent Baselines

45 citations · 138 across the 16 of their papers we have counts for

collaborators
Showing 2021Show all

4 papers · 1 filter