7 citations · 7 across the 1 of their papers we have counts for
5 papers · 1 filter
Task-Optimal Exploration in Linear Dynamical Systems
Andrew Wagenmaker, Max Simchowitz, Kevin Jamieson
Exploration in unknown environments is a fundamental problem in reinforcement learning and control. In this work, we study task-guided exploration and determine what precisely an a…
Non-Asymptotic Gap-Dependent Regret Bounds for Tabular MDPs
Max Simchowitz, Kevin Jamieson
This paper establishes that optimistic algorithms attain gap-dependent and non-asymptotic logarithmic regret for episodic MDPs. In contrast to prior work, our bounds do not suffer…
Adaptive Sampling for Convex Regression
Max Simchowitz, Kevin Jamieson, Jordan W. Suchow +1
In this paper, we introduce the first principled adaptive-sampling procedure for learning a convex function in the norm, a problem that arises often in the behavioral an…
On the Randomized Complexity of Minimizing a Convex Quadratic Function
Max Simchowitz
Minimizing a convex, quadratic objective of the form for is a…
Approximate Ranking from Pairwise Comparisons
Reinhard Heckel, Max Simchowitz, Kannan Ramchandran +1
A common problem in machine learning is to rank a set of n items based on pairwise comparisons. Here ranking refers to partitioning the items into sets of pre-specified sizes accor…