◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Will Dabney

26 papers hereh-index 3411.7k citations64 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author16
  • last author6

Across the 26 of 26 papers where every author was matched, so the position is known.

fields
  • cs.LG20
  • cs.AI4
  • stat.ML2
same name
  • Will Dabney — 6 papers
  • Will Dabney — 3 papers, h 6
  • Will Dabney — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172022
most citedRainbow: Combining Improvements in Deep Reinforcement Learning

424 citations · 1.3k across the 16 of their papers we have counts for

collaborators
Showing 2020Show all

4 papers · 1 filter

cs.LG2020★ 81 cited

Revisiting Fundamentals of Experience Replay

William Fedus, Prajit Ramachandran, Rishabh Agarwal +4

Experience replay is central to off-policy algorithms in deep reinforcement learning (RL), but there remain significant gaps in our understanding. We therefore present a systematic…

cs.AI2020

Deep Reinforcement Learning and its Neuroscientific Implications

Matthew Botvinick, Jane X. Wang, Will Dabney +2

The emergence of powerful artificial intelligence is defining new research directions in neuroscience. To date, this research has focused largely on deep neural networks trained us…

cs.LG2020★ 2 cited

Temporally-Extended ε-Greedy Exploration

Will Dabney, Georg Ostrovski, André Barreto

Recent work on exploration in reinforcement learning (RL) has led to a series of increasingly complex solutions to the problem. This increase in complexity often comes at the expen…

cs.LG2020

The Value-Improvement Path: Towards Better Representations for Reinforcement Learning

Will Dabney, André Barreto, Mark Rowland +4

In value-based reinforcement learning (RL), unlike in supervised learning, the agent faces not a single, stationary, approximation problem, but a sequence of value prediction probl…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.