91 citations · 355 across the 18 of their papers we have counts for
Showing 2019 · stat.MLShow all
2 papers · 2 filters
stat.ML2019★ 53 cited
Optimism in Reinforcement Learning with Generalized Linear Function Approximation
Yining Wang, Ruosong Wang, Simon S. Du +1
We design a new provably efficient algorithm for episodic reinforcement learning with generalized linear function approximation. We analyze the algorithm under a new expressivity a…
stat.ML2019
Contextual Bandits with Continuous Actions: Smoothing, Zooming, and Adapting
Akshay Krishnamurthy, John Langford, Aleksandrs Slivkins +1
We study contextual bandit learning with an abstract policy class and continuous action space. We obtain two qualitatively different regret bounds: one competes with a smoothed ver…