3 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.LG2024
OPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple Estimators
Allen Nie, Yash Chandak, Christina J. Yuan +3
Offline policy evaluation (OPE) allows us to evaluate and estimate a new sequential decision-making policy's performance by leveraging historical interaction data collected from ot…
stat.ME2024★ 2 cited
Minimax-Regret Sample Selection in Randomized Experiments
Yuchen Hu, Henry Zhu, Emma Brunskill +1
Randomized controlled trials are often run in settings with many subpopulations that may have differential benefits from the treatment being evaluated. We consider the problem of s…
cs.LG2023★ 3 cited
Off-Policy Evaluation for Action-Dependent Non-Stationary Environments
Yash Chandak, Shiv Shankar, Nathaniel D. Bastian +3
Methods for sequential decision-making are often built upon a foundational assumption that the underlying decision process is stationary. This limits the application of such method…