activity
20162021
most citedPolicy Gradient Methods for Reinforcement Learning with Function Approximation and Action-Dependent Baselines

45 citations · 138 across the 16 of their papers we have counts for

collaborators
Showing 2016 · math.STShow all

Nothing from them under that filter.

Their other years and fields are still on the left.