◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Rémi Munos

8 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author5

Across the 7 of 8 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • math.ST2
  • cs.AI1
  • stat.ML1
same name
  • Rémi Munos — 22 papers, h 14
  • Rémi Munos — 6 papers, h 2
  • Rémi Munos — 3 papers
  • Rémi Munos — 1 paper, h 2
  • Rémi Munos — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20122022
most citedBandit Algorithms for Tree Search

178 citations · 427 across the 8 of their papers we have counts for

collaborators
Showing 2012Show all

4 papers · 1 filter

cs.LG2012★ 42 cited

On the Sample Complexity of Reinforcement Learning with a Generative Model

Mohammad Gheshlaghi Azar, Remi Munos, Bert Kappen

We consider the problem of learning the optimal action-value function in the discounted-reward Markov decision processes (MDPs). We prove a new PAC bound on the sample-complexity o…

math.ST2012

Minimax Number of Strata for Online Stratified Sampling given Noisy Samples

Alexandra Carpentier, Rémi Munos

We consider the problem of online stratified sampling for Monte Carlo integration of a function given a finite budget of n noisy evaluations to the function. More precisely we fo…

math.ST2012

Bandit Theory meets Compressed Sensing for high dimensional Stochastic Linear Bandit

Alexandra Carpentier, Rémi Munos

We consider a linear stochastic bandit problem where the dimension K of the unknown parameter θ is larger than the sampling budget n. In such cases, it is in general impossib…

stat.ML2012★ 36 cited

Thompson Sampling: An Asymptotically Optimal Finite Time Analysis

Emilie Kaufmann, Nathaniel Korda, Rémi Munos

The question of the optimality of Thompson Sampling for solving the stochastic multi-armed bandit problem had been open since 1933. In this paper we answer it positively for the ca…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.