Evaluating Causal Models by Comparing Interventional Distributions
arXiv:1608.04698
Abstract
The predominant method for evaluating the quality of causal models is to measure the graphical accuracy of the learned model structure. We present an alternative method for evaluating causal models that directly measures the accuracy of estimated interventional distributions. We contrast such distributional measures with structural measures, such as structural Hamming distance and structural intervention distance, showing that structural measures often correspond poorly to the accuracy of estimated interventional distributions. We use a number of real and synthetic datasets to illustrate various scenarios in which structural measures provide misleading results with respect to algorithm selection and parameter tuning, and we recommend that distributional measures become the new standard for evaluating causal models.
References in corpus (6)
- Searching for Bayesian Network Structures in the Space of Restricted Acyclic Partially Directed Graphs
- Finding Optimal Bayesian Networks
- On the Validity of Covariate Adjustment for Estimating Causal Effects
- Bayesian structure learning using dynamic programming and MCMC
- Adjacency-Faithfulness and Conservative Causal Inference
- Testing Identifiability of Causal Effects