activity
20122022
most citedPolicy Gradient Methods for Reinforcement Learning with Function Approximation and Action-Dependent Baselines

45 citations · 215 across the 20 of their papers we have counts for

collaborators
Showing 2016 · cs.AIShow all

Nothing from them under that filter.

Their other years and fields are still on the left.