20 citations · 31 across the 3 of their papers we have counts for
3 papers
cs.LG2021★ 20 cited
Accelerating Quadratic Optimization with Reinforcement Learning
Jeffrey Ichnowski, Paras Jain, Bartolomeo Stellato +6
First-order methods for quadratic optimization such as OSQP are widely used for large-scale machine learning and embedded optimal control, where many related problems must be rapid…
cs.LG2019★ 1 cited
Hierarchical Variational Imitation Learning of Control Programs
Roy Fox, Richard Shin, William Paul +5
Autonomous agents can learn by imitating teacher demonstrations of the intended behavior. Hierarchical control policies are ubiquitously useful for such learning, having the potent…
cs.RO2017★ 10 cited
DDCO: Discovery of Deep Continuous Options for Robot Learning from Demonstrations
Sanjay Krishnan, Roy Fox, Ion Stoica +1
An option is a short-term skill consisting of a control policy for a specified region of the state space, and a termination condition recognizing leaving that region. In prior work…