◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

B. Scherrer

5 papers hereh-index 262.2k citations67 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author2
  • first author1
  • last author2

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.LG1
  • math.OC1

identity via Semantic Scholar / OpenAlex

activity
20062013
most citedShould one compute the Temporal Difference fix point or minimize the Bellman Residual? The unified oblique projection view

44 citations · 100 across the 5 of their papers we have counts for

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2013★ 39 cited

Off-policy Learning with Eligibility Traces: A Survey

Matthieu Geist, Bruno Scherrer

In the framework of Markov Decision Processes, off-policy learning, that is the problem of learning a linear approximation of the value function of some fixed policy from one traje…

cs.AI2010★ 44 cited

Should one compute the Temporal Difference fix point or minimize the Bellman Residual? The unified oblique projection view

Bruno Scherrer

We investigate projection methods, for evaluating a linear approximation of the value function of a policy in a Markov Decision Process context. We consider two popular approaches,…

cs.AI2006

Modular self-organization

Bruno Scherrer

The aim of this paper is to provide a sound framework for addressing a difficult problem: the automatic construction of an autonomous agent's modular architecture. We combine resul…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.