◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Will Dabney

26 papers hereh-index 3411.7k citations64 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author16
  • last author6

Across the 26 of 26 papers where every author was matched, so the position is known.

fields
  • cs.LG20
  • cs.AI4
  • stat.ML2
same name
  • Will Dabney — 7 papers
  • Will Dabney — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172022
most citedRainbow: Combining Improvements in Deep Reinforcement Learning

424 citations · 1.3k across the 16 of their papers we have counts for

collaborators
Showing 2017Show all

4 papers · 1 filter

cs.AI2017★ 149 cited

Distributional Reinforcement Learning with Quantile Regression

Will Dabney, Mark Rowland, Marc G. Bellemare +1

In reinforcement learning an agent interacts with the environment by taking actions and observing the next state and reward. When sampled probabilistically, these state transitions…

cs.AI2017★ 424 cited

Rainbow: Combining Improvements in Deep Reinforcement Learning

Matteo Hessel, Joseph Modayil, Hado van Hasselt +7

The deep reinforcement learning community has made several independent improvements to the DQN algorithm. However, it is unclear which of these extensions are complementary and can…

cs.LG2017★ 241 cited

A Distributional Perspective on Reinforcement Learning

Marc G. Bellemare, Will Dabney, Rémi Munos

In this paper we argue for the fundamental importance of the value distribution: the distribution of the random return received by a reinforcement learning agent. This is in contra…

cs.LG2017★ 256 cited

The Cramer Distance as a Solution to Biased Wasserstein Gradients

Marc G. Bellemare, Ivo Danihelka, Will Dabney +4

The Wasserstein probability metric has received much attention from the machine learning community. Unlike the Kullback-Leibler divergence, which strictly measures change in probab…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.