Assessing lack of common support in causal inference using Bayesian nonparametrics: Implications for evaluating the effect of breastfeeding on children's cognitive outcomes
arXiv:1311.7244 · doi:10.1214/13-AOAS630
Abstract
Causal inference in observational studies typically requires making comparisons between groups that are dissimilar. For instance, researchers investigating the role of a prolonged duration of breastfeeding on child outcomes may be forced to make comparisons between women with substantially different characteristics on average. In the extreme there may exist neighborhoods of the covariate space where there are not sufficient numbers of both groups of women (those who breastfed for prolonged periods and those who did not) to make inferences about those women. This is referred to as lack of common support. Problems can arise when we try to estimate causal effects for units that lack common support, thus we may want to avoid inference for such units. If ignorability is satisfied with respect to a set of potential confounders, then identifying whether, or for which units, the common support assumption holds is an empirical question. However, in the high-dimensional covariate space often required to satisfy ignorability such identification may not be trivial. Existing methods used to address this problem often require reliance on parametric assumptions and most, if not all, ignore the information embedded in the response variable. We distinguish between the concepts of "common support" and common causal support." We propose a new approach for identifying common causal support that addresses some of the shortcomings of existing methods. We motivate and illustrate the approach using data from the National Longitudinal Survey of Youth to estimate the effect of breastfeeding at least nine months on reading and math achievement scores at age five or six. We also evaluate the comparative performance of this method in hypothetical examples and simulations where the true treatment effect is known.
Published in at http://dx.doi.org/10.1214/13-AOAS630 the Annals of Applied Statistics (http://www.imstat.org/aoas/) by the Institute of Mathematical Statistics (http://www.imstat.org)
Cited by in corpus (15)
- Discussion on "Bayesian Regression Tree Models for Causal Inference: Regularization, Confounding, and Heterogeneous Effects" by Hahn, Murray and Carvalho
- Matching for balance, pairing for heterogeneity in an observational study of the effectiveness of for-profit and not-for-profit high schools in Chile
- Response Transformation and Profit Decomposition for Revenue Uplift Modeling
- Automated versus do-it-yourself methods for causal inference: Lessons learned from a data analysis competition
- Identifying Causal-Effect Inference Failure with Uncertainty-Aware Models
- An Evaluation Toolkit to Guide Model Selection and Cohort Definition in Causal Inference
- Estimation of Causal Effects of Multiple Treatments in Observational Studies with a Binary Outcome
- From controlled to undisciplined data: estimating causal effects in the era of data science using a potential outcome framework
- Observational-Interventional Priors for Dose-Response Learning
- Sequential Deconfounding for Causal Inference with Unobserved Confounders
- Machine learning in the social and health sciences
- Generalizing Off-Policy Learning under Sample Selection Bias
- Causal inference and machine learning approaches for evaluation of the health impacts of large-scale air quality regulations
- A causal fused lasso for interpretable heterogeneous treatment effects estimation
- Modelling hetegeneous treatment effects by quantitle local polynomial decision tree and forest