7 citations · 12 across the 2 of their papers we have counts for
4 papers
Adapting to Delays and Data in Adversarial Multi-Armed Bandits
András György, Pooria Joulani
We consider the adversarial multi-armed bandit problem under delayed feedback. We analyze variants of the Exp3 algorithm that tune their step-size using only information (about the…
Adaptive Approximate Policy Iteration
Botao Hao, Nevena Lazic, Yasin Abbasi-Yadkori +2
Model-free reinforcement learning algorithms combined with value function approximation have recently achieved impressive performance in a variety of application domains. However,…
A Modular Analysis of Adaptive (Non-)Convex Optimization: Optimism, Composite Objectives, and Variational Bounds
Pooria Joulani, András György, Csaba Szepesvári
Recently, much work has been done on extending the scope of online learning and incremental stochastic optimization algorithms. In this paper we contribute to this effort in two wa…
Fast Cross-Validation for Incremental Learning
Pooria Joulani, András György, Csaba Szepesvári
Cross-validation (CV) is one of the main tools for performance estimation and parameter tuning in machine learning. The general recipe for computing CV estimate is to run a learnin…