702 citations
- Google DeepMind (United Kingdom)GB54 papers
- Google (United States)US22 papers
- Massachusetts Institute of TechnologyUS11 papers
- University College LondonGB8 papers
- ETH ZurichCH6 papers
- Stanford UniversityUS6 papers
- École Normale Supérieure - PSLFR5 papers
- École PolytechniqueFR5 papers
- Imperial College LondonGB5 papers
- Institut national de recherche en sciences et technologies du numériqueFR5 papers
- Carnegie Mellon UniversityUS4 papers
- Centre de Mathématiques Appliquées de l'École polytechniqueFR4 papers
Showing 2022 · stat.MLShow all
2 papers · 2 filters
stat.ML2022★ 3 cited
Curiosity in Hindsight: Intrinsic Exploration in Stochastic Environments
Daniel Jarrett, Corentin Tallec, Florent Altché +3
Consider the problem of exploration in sparse-reward or reward-free environments, such as in Montezuma's Revenge. In the curiosity-driven paradigm, the agent is rewarded for how mu…
stat.ML2022★ 3 cited
From Dirichlet to Rubin: Optimistic Exploration in RL without Bonuses
Daniil Tiapkin, Denis Belomestny, Eric Moulines +5
We propose the Bayes-UCBVI algorithm for reinforcement learning in tabular, stage-dependent, episodic Markov decision process: a natural extension of the Bayes-UCB algorithm by Kau…