◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

B. Scherrer

13 papers hereh-index 262.2k citations67 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author2
  • first author1
  • middle author7
  • last author3

Across the 13 of 13 papers where every author was matched, so the position is known.

fields
  • cs.LG8
  • cs.AI4
  • math.OC1
same name
  • B. Scherrer — 2 papers
  • B. Scherrer — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20062021
most citedA Theory of Regularized Markov Decision Processes

90 citations · 194 across the 7 of their papers we have counts for

collaborators
Showing 2018Show all

4 papers · 1 filter

cs.LG2018

Anderson Acceleration for Reinforcement Learning

Matthieu Geist, Bruno Scherrer

Anderson acceleration is an old and simple method for accelerating the computation of a fixed point. However, as far as we know and quite surprisingly, it has never been applied to…

cs.LG2018

How to Combine Tree-Search Methods in Reinforcement Learning

Yonathan Efroni, Gal Dalal, Bruno Scherrer +1

Finite-horizon lookahead policies are abundantly used in Reinforcement Learning and demonstrate impressive empirical success. Usually, the lookahead policies are implemented with s…

cs.LG2018

Multiple-Step Greedy Policies in Online and Approximate Reinforcement Learning

Yonathan Efroni, Gal Dalal, Bruno Scherrer +1

Multiple-step lookahead policies have demonstrated high empirical competence in Reinforcement Learning, via the use of Monte Carlo Tree Search or Model Predictive Control. In a rec…

cs.AI2018

Beyond the One Step Greedy Approach in Reinforcement Learning

Yonathan Efroni, Gal Dalal, Bruno Scherrer +1

The famous Policy Iteration algorithm alternates between policy improvement and policy evaluation. Implementations of this algorithm with several variants of the latter evaluation…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.