2 citations · 3 across the 2 of their papers we have counts for
2 papers
stat.ML2022★ 1 cited
Instance-Dependent Confidence and Early Stopping for Reinforcement Learning
Koulik Khamaru, Eric Xia, Martin J. Wainwright +1
Various algorithms for reinforcement learning (RL) exhibit dramatic variation in their convergence rates as a function of problem structure. Such problem-dependent behavior is not…
stat.ML2021★ 2 cited
Instance-optimality in optimal value estimation: Adaptivity via variance-reduced Q-learning
Koulik Khamaru, Eric Xia, Martin J. Wainwright +1
Various algorithms in reinforcement learning exhibit dramatic variability in their convergence rates and ultimate accuracy as a function of the problem structure. Such instance-spe…