299 citations · 478 across the 3 of their papers we have counts for
3 papers
cs.LG2011★ 119 cited
Efficient Optimal Learning for Contextual Bandits
Miroslav Dudik, Daniel Hsu, Satyen Kale +4
We address the problem of learning in an online setting where the learner repeatedly observes features, selects among a set of actions, and receives reward for the action taken. We…
cs.LG2011★ 299 cited
Doubly Robust Policy Evaluation and Learning
Miroslav Dudik, John Langford, Lihong Li
We study decision making in environments where the reward is only partially observed, but can be modeled as a function of an action and an observed context. This setting, known as…
q-bio.QM2007★ 60 cited
Faster solutions of the inverse pairwise Ising problem
Tamara Broderick, Miroslav Dudik, Gasper Tkacik +2
Recent work has shown that probabilistic models based on pairwise interactions-in the simplest case, the Ising model-provide surprisingly accurate descriptions of experiments on re…