4 papers
FUSE: Ensembling Verifiers with Zero Labeled Data
Joonhyuk Lee, Virginia Ma, Sarah Zhao +4
Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In practice, this often involves u…
Using Individualized Treatment Effects to Assess Treatment Effect Heterogeneity
Konstantinos Sechidis, Cong Zhang, Sophie Sun +3
Assessing treatment effect heterogeneity (TEH) in clinical trials is crucial, as it provides insights into the variability of treatment responses among patients, influencing import…
Chiseling: Powerful and Valid Subgroup Selection via Interactive Machine Learning
Nathan Cheng, Asher Spector, Lucas Janson
In regression and causal inference, controlled subgroup selection aims to identify, with inferential guarantees, a subgroup (defined as a subset of the covariate space) on which th…
Mosaic inference on panel data
Asher Spector, Rina Foygel Barber, Emmanuel Candès
Analysis of panel data via linear regression is widespread across disciplines. To perform statistical inference, such analyses typically assume that clusters of observations are jo…