27 citations · 127 across the 31 of their papers we have counts for
Showing eess.SYShow all
2 papers · 1 filter
eess.SY2019★ 24 cited
MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding
Wenbo Zhang, Osbert Bastani, Vijay Kumar
Reinforcement learning is a promising approach to learning control policies for performing complex multi-agent robotics tasks. However, a policy learned in simulation often fails t…
eess.SY2019★ 7 cited
Robust Model Predictive Shielding for Safe Reinforcement Learning with Stochastic Dynamics
Shuo Li, Osbert Bastani
This paper proposes a framework for safe reinforcement learning that can handle stochastic nonlinear dynamical systems. We focus on the setting where the nominal dynamics are known…