133 citations · 322 across the 19 of their papers we have counts for
3 papers · 1 filter
Differentiable MPC for End-to-end Planning and Control
Brandon Amos, Ivan Dario Jimenez Rodriguez, Jacob Sacks +2
We present foundations for using Model Predictive Control (MPC) as a differentiable policy class for reinforcement learning in continuous state and action spaces. This provides one…
Depth-Limited Solving for Imperfect-Information Games
Noam Brown, Tuomas Sandholm, Brandon Amos
A fundamental challenge in imperfect-information games is that states do not have well-defined values. As a result, depth-limited search algorithms used in single-agent settings an…
Learning Awareness Models
Brandon Amos, Laurent Dinh, Serkan Cabi +7
We consider the setting of an agent with a fixed body interacting with an unknown and uncertain external world. We show that models trained to predict proprioceptive information ab…