37 citations · 141 across the 59 of their papers we have counts for
Showing 2017 · stat.MLShow all
2 papers · 2 filters
stat.ML2017
Instrument-Armed Bandits
Nathan Kallus
We extend the classic multi-armed bandit (MAB) model to the setting of noncompliance, where the arm pull is a mere instrument and the treatment applied may differ from it, which gi…
stat.ML2017
Balanced Policy Evaluation and Learning
Nathan Kallus
We present a new approach to the problems of evaluating and learning personalized decision policies from observational data of past contexts, decisions, and outcomes. Only the outc…