82 citations · 341 across the 47 of their papers we have counts for
4 papers · 1 filter
Model-Based Quality-Diversity Search for Efficient Robot Learning
Leon Keller, Daniel Tanneberg, Svenja Stark +1
Despite recent progress in robot learning, it still remains a challenge to program a robot to deal with open-ended object manipulation tasks. One approach that was recently used to…
Generalized Mean Estimation in Monte-Carlo Tree Search
Tuan Dam, Pascal Klink, Carlo D'Eramo +2
We consider Monte-Carlo Tree Search (MCTS) applied to Markov Decision Processes (MDPs) and Partially Observable MDPs (POMDPs), and the well-known Upper Confidence bound for Trees (…
Information Gathering in Decentralized POMDPs by Policy Graph Improvement
Mikko Lauri, Joni Pajarinen, Jan Peters
Decentralized policies for information gathering are required when multiple autonomous agents are deployed to collect data about a phenomenon of interest without the ability to com…
Intrinsic Motivation and Mental Replay enable Efficient Online Adaptation in Stochastic Recurrent Networks
Daniel Tanneberg, Jan Peters, Elmar Rueckert
Autonomous robots need to interact with unknown, unstructured and changing environments, constantly facing novel challenges. Therefore, continuous online adaptation for lifelong-le…