201 citations · 201 across the 1 of their papers we have counts for
2 papers
cs.LG2017
Safe Exploration for Identifying Linear Systems via Robust Optimization
Tyler Lu, Martin Zinkevich, Craig Boutilier +2
Safely exploring an unknown dynamical system is critical to the deployment of reinforcement learning (RL) in physical systems where failures may have catastrophic consequences. In…
math.OC2009★ 201 cited
Slow Learners are Fast
John Langford, Alexander Smola, Martin Zinkevich
Online learning algorithms have impressive convergence properties when it comes to risk minimization and convex games on very large problems. However, they are inherently sequentia…