101 citations · 225 across the 14 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2016★ 67 cited
Safe Policy Improvement by Minimizing Robust Baseline Regret
Marek Petrik, Yinlam Chow, Mohammad Ghavamzadeh
An important problem in sequential decision-making under uncertainty is to use limited data to compute a safe policy, i.e., a policy that is guaranteed to perform at least as well…
stat.ML2016
Building an Interpretable Recommender via Loss-Preserving Transformation
Amit Dhurandhar, Sechan Oh, Marek Petrik
We propose a method for building an interpretable recommender system for personalizing online content and promotions. Historical data available for the system consists of customer…