Related Experiment Video
Updated: Sep 28, 2025

Inverse Probability of Treatment Weighting Propensity Score using the Military Health System Data Repository and National Death Index
Published on: January 8, 2020
On the Use of Covariate Supersets for Identification Conditions
Paul N Zivich1, Bonnie E Shook-Sa2, Jessie K Edwards1
1From the Department of Epidemiology, UNC Gillings School of Global Public Health, Chapel Hill, NC.
Abstract:
The union of distinct covariate sets, or the superset, is often used in proofs for the identification or the statistical consistency of an estimator when multiple sources of bias are present. However, the use of a superset can obscure important nuances. Here, we provide two illustrative examples: one in the context of missing data on outcomes, and one in which the average causal effect is transported to another target population. As these examples demonstrate, the use of supersets may indicate a parameter is not identifiable when the parameter is indeed identified. Furthermore, a series of exchangeability conditions may lead to successively weaker conditions. Future work on approaches to address multiple biases can avoid these pitfalls by considering the more general case of nonoverlapping covariate sets.
Related Concept Videos
Comparing the Survival Analysis of Two or More Groups
Confounding in Epidemiological Studies
Contingency Table
Study Design in Statistics
Does aspirin reduce the risk of heart attacks? Is one brand of fertilizer more effective at growing roses than another? Is fatigue as dangerous to a driver as the influence of alcohol? Questions like these are answered using randomized experiments with proper...
Assumptions of Survival Analysis
Factorial Design

