58 citations · 107 across the 9 of their papers we have counts for
4 papers · 1 filter
Learning to Plan via a Multi-Step Policy Regression Method
Stefan Wagner, Michael Janschek, Tobias Uelwer +1
We propose a new approach to increase inference performance in environments that require a specific sequence of actions in order to be solved. This is for example the case for maze…
Smaller World Models for Reinforcement Learning
Jan Robine, Tobias Uelwer, Stefan Harmeling
Sample efficiency remains a fundamental issue of reinforcement learning. Model-based algorithms try to make better use of data by simulating the environment with a model. We propos…
On the Vulnerability of Capsule Networks to Adversarial Attacks
Felix Michels, Tobias Uelwer, Eric Upschulte +1
This paper extensively evaluates the vulnerability of capsule networks to different adversarial attacks. Recent work suggests that these architectures are more robust towards adver…
Modular Block-diagonal Curvature Approximations for Feedforward Architectures
Felix Dangel, Stefan Harmeling, Philipp Hennig
We propose a modular extension of backpropagation for the computation of block-diagonal approximations to various curvature matrices of the training objective (in particular, the H…