20 citations · 62 across the 28 of their papers we have counts for
Showing 2019 · stat.MLShow all
2 papers · 2 filters
stat.ML2019
Instance-dependent -bounds for policy evaluation in tabular reinforcement learning
Ashwin Pananjady, Martin J. Wainwright
Markov reward processes (MRPs) are used to model stochastic phenomena arising in operations research, control engineering, robotics, and artificial intelligence, as well as communi…
stat.ML2019★ 13 cited
Max-Affine Regression: Provable, Tractable, and Near-Optimal Statistical Estimation
Avishek Ghosh, Ashwin Pananjady, Adityanand Guntuboyina +1
Max-affine regression refers to a model where the unknown regression function is modeled as a maximum of unknown affine functions for a fixed . This generalizes linea…