◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Schuurmans

23 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author8
  • last author13

Across the 21 of 23 papers where every author was matched, so the position is known.

fields
  • cs.LG17
  • cs.AI3
  • stat.ML2
  • cs.CL1

identity via Semantic Scholar / OpenAlex

activity
20182025
most citedGenDICE: Generalized Offline Estimation of Stationary Values

49 citations · 121 across the 17 of their papers we have counts for

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2021★ 5 cited

Characterizing the Gap Between Actor-Critic and Policy Gradient

Junfeng Wen, Saurabh Kumar, Ramki Gummadi +1

Actor-critic (AC) methods are ubiquitous in reinforcement learning. Although it is understood that AC methods are closely related to policy gradient (PG), their precise connection…

cs.AI2021

Joint Attention for Multi-Agent Coordination and Social Learning

Dennis Lee, Natasha Jaques, Chase Kew +4

Joint attention - the ability to purposefully coordinate attention with another agent, and mutually attend to the same thing -- is a critical component of human social cognition. I…

cs.AI2018

Planning and Learning with Stochastic Action Sets

Craig Boutilier, Alon Cohen, Amit Daniely +5

In many practical uses of reinforcement learning (RL) the set of actions available at a given state is a random variable, with realizations governed by an exogenous stochastic proc…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.