44 citations · 100 across the 5 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2013★ 39 cited
Off-policy Learning with Eligibility Traces: A Survey
Matthieu Geist, Bruno Scherrer
In the framework of Markov Decision Processes, off-policy learning, that is the problem of learning a linear approximation of the value function of some fixed policy from one traje…
cs.AI2010★ 44 cited
Should one compute the Temporal Difference fix point or minimize the Bellman Residual? The unified oblique projection view
Bruno Scherrer
We investigate projection methods, for evaluating a linear approximation of the value function of a policy in a Markov Decision Process context. We consider two popular approaches,…
cs.AI2006
Modular self-organization
Bruno Scherrer
The aim of this paper is to provide a sound framework for addressing a difficult problem: the automatic construction of an autonomous agent's modular architecture. We combine resul…