activity
20122021
most citedDistral: Robust Multitask Reinforcement Learning

182 citations · 420 across the 9 of their papers we have counts for

collaborators
Showing cs.LGShow all

8 papers · 1 filter

cs.LG202016 cited

Combining Q-Learning and Search with Amortized Value Estimates

Jessica B. Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez +4

We introduce "Search with Amortized Value Estimates" (SAVE), an approach for combining model-free Q-learning with model-based Monte-Carlo Tree Search (MCTS). In SAVE, a learned pri…

cs.LG20191 cited

Object-oriented state editing for HRL

Victor Bapst, Alvaro Sanchez-Gonzalez, Omar Shams +4

We introduce agents that use object-oriented reasoning to consider alternate states of the world in order to more quickly find solutions to problems. Specifically, a hierarchical c…

cs.LG201976 cited

Hamiltonian Graph Networks with ODE Integrators

Alvaro Sanchez-Gonzalez, Victor Bapst, Kyle Cranmer +1

We introduce an approach for imposing physically informed inductive biases in learned simulation models. We combine graph networks with a differentiable ordinary differential equat…

cs.LG20192 cited

Structured agents for physical construction

Victor Bapst, Alvaro Sanchez-Gonzalez, Carl Doersch +4

Physical construction---the ability to compose objects, subject to physical dynamics, to serve some function---is fundamental to human intelligence. We introduce a suite of challen…

cs.LG2018

Relational Deep Reinforcement Learning

Vinicius Zambaldi, David Raposo, Adam Santoro +13

We introduce an approach for deep reinforcement learning (RL) that improves upon the efficiency, generalization capacity, and interpretability of conventional approaches through st…

cs.LG2018

Relational inductive bias for physical construction in humans and machines

Jessica B. Hamrick, Kelsey R. Allen, Victor Bapst +4

While current deep learning systems excel at tasks such as object classification, language processing, and gameplay, few can construct or modify a complex system such as a tower of…