91 citations · 91 across the 1 of their papers we have counts for
4 papers
Improved Regret Bounds for Oracle-Based Adversarial Contextual Bandits
Vasilis Syrgkanis, Haipeng Luo, Akshay Krishnamurthy +1
We give an oracle-based algorithm for the adversarial contextual bandit problem, where either contexts are drawn i.i.d. or the sequence of contexts is known a priori, but where the…
Exploratory Gradient Boosting for Reinforcement Learning in Complex Domains
David Abel, Alekh Agarwal, Fernando Diaz +2
High-dimensional observations and complex real-world dynamics present major challenges in reinforcement learning for both function approximation and exploration. We address both of…
Learning to Search Better Than Your Teacher
Kai-Wei Chang, Akshay Krishnamurthy, Alekh Agarwal +2
Methods for learning to search for structured prediction typically imitate a reference policy, with existing theoretical guarantees demonstrating low regret compared to that refere…
Detecting Activations over Graphs using Spanning Tree Wavelet Bases
James Sharpnack, Akshay Krishnamurthy, Aarti Singh
We consider the detection of activations over graphs under Gaussian noise, where signals are piece-wise constant over the graph. Despite the wide applicability of such a detection…