153 citations · 362 across the 29 of their papers we have counts for
Showing 2016 · stat.MLShow all
2 papers · 2 filters
stat.ML2016
Causal Bandits: Learning Good Interventions via Causal Inference
Finnian Lattimore, Tor Lattimore, Mark D. Reid
We study the problem of using causal models to improve the rate at which good interventions can be learned online in a stochastic environment. Our formalism combines multi-arm band…
stat.ML2016
Conservative Bandits
Yifan Wu, Roshan Shariff, Tor Lattimore +1
We study a novel multi-armed bandit problem that models the challenge faced by a company wishing to explore new strategies to maximize revenue whilst simultaneously maintaining the…