2 citations · 4 across the 3 of their papers we have counts for
4 papers
Online Policy Gradient for Model Free Learning of Linear Quadratic Regulators with Regret
Asaf Cassel, Tomer Koren
We consider the task of learning to control a linear dynamical system under fixed quadratic costs, known as the Linear Quadratic Regulator (LQR) problem. While model-free approache…
The Pendulum Arrangement: Maximizing the Escape Time of Heterogeneous Random Walks
Asaf Cassel, Shie Mannor, Guy Tennenholtz
We identify a fundamental phenomenon of heterogeneous one dimensional random walks: the escape (traversal) time is maximized when the heterogeneity in transition probabilities form…
Bandit Linear Control
Asaf Cassel, Tomer Koren
We consider the problem of controlling a known linear dynamical system under stochastic noise, adversarially chosen costs, and bandit feedback. Unlike the full feedback setting whe…
Logarithmic Regret for Learning Linear Quadratic Regulators Efficiently
Asaf Cassel, Alon Cohen, Tomer Koren
We consider the problem of learning in Linear Quadratic Control systems whose transition parameters are initially unknown. Recent results in this setting have demonstrated efficien…