◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Marc G. Bellemare

41 papers hereh-index 4450.2k citations92 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author19
  • last author17

Across the 40 of 41 papers where every author was matched, so the position is known.

fields
  • cs.LG30
  • cs.AI4
  • stat.ML3
  • cond-mat.mes-hall1
  • cs.CL1
  • cs.HC1
same name
  • Marc G. Bellemare — 3 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172026
most citedThe Cramer Distance as a Solution to Biased Wasserstein Gradients

256 citations · 1.1k across the 23 of their papers we have counts for

collaborators
Showing cs.AIShow all

4 papers · 1 filter

cs.AI2020

The Importance of Pessimism in Fixed-Dataset Policy Optimization

Jacob Buckman, Carles Gelada, Marc G. Bellemare

We study worst-case guarantees on the expected return of fixed-dataset policy optimization algorithms. Our core contribution is a unified conceptual and mathematical framework for…

cs.AI2019

Algorithmic Improvements for Deep Reinforcement Learning applied to Interactive Fiction

Vishal Jain, William Fedus, Hugo Larochelle +2

Text-based games are a natural challenge domain for deep reinforcement learning algorithms. Their state and action spaces are combinatorially large, their reward function is sparse…

cs.AI2018

The Barbados 2018 List of Open Issues in Continual Learning

Tom Schaul, Hado van Hasselt, Joseph Modayil +7

We want to make progress toward artificial general intelligence, namely general-purpose agents that autonomously learn how to competently act in complex environments. The purpose o…

cs.AI2017★ 149 cited

Distributional Reinforcement Learning with Quantile Regression

Will Dabney, Mark Rowland, Marc G. Bellemare +1

In reinforcement learning an agent interacts with the environment by taking actions and observing the next state and reward. When sampled probabilistically, these state transitions…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.