31 citations · 40 across the 2 of their papers we have counts for
4 papers
CausalWorld: A Robotic Manipulation Benchmark for Causal Structure and Transfer Learning
Ossama Ahmed, Frederik Träuble, Anirudh Goyal +5
Despite recent successes of reinforcement learning (RL), it remains a challenge for agents to transfer learned skills to related environments. To facilitate research addressing thi…
Learning explanations that are hard to vary
Giambattista Parascandolo, Alexander Neitz, Antonio Orvieto +2
In this paper, we investigate the principle that `good explanations are hard to vary' in the context of deep learning. We show that averaging gradients across examples -- akin to a…
Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning
Giambattista Parascandolo, Lars Buesing, Josh Merel +6
Standard planners for sequential decision making (including Monte Carlo planning, tree search, dynamic programming, etc.) are constrained by an implicit sequential planning assumpt…
Adaptive Skip Intervals: Temporal Abstraction for Recurrent Dynamical Models
Alexander Neitz, Giambattista Parascandolo, Stefan Bauer +1
We introduce a method which enables a recurrent dynamics model to be temporally abstract. Our approach, which we call Adaptive Skip Intervals (ASI), is based on the observation tha…