Related Experiment Video
Updated: Dec 30, 2025

Inverse Probability of Treatment Weighting Propensity Score using the Military Health System Data Repository and National Death Index
Published on: January 8, 2020
Improving measurement of binary covariates in claims data: A simulation study
John G Connolly1, Robert J Glynn1, Sebastian Schneeweiss1
1Division of Pharmacoepidemiology and Pharmacoeconomics, Department of Medicine, Brigham and Women's Hospital and Harvard Medical School, Harvard University, Boston, Massachusetts.
Purpose:
When investigators have two claims-based definitions for a binary confounder, it is unclear whether to prefer the more sensitive or more specific definition. Our objective was to compare adjusting for the sensitive or specific definition alone vs two novel approaches combining both definitions: a "two-algorithm indicator" and a "two-algorithm restriction" approach.
Methods:
Each simulated patient had a binary exposure, outcome, and confounder. We created two nested, misclassified versions of the confounder using validated heart failure definitions. The sensitive definition had a sensitivity/specificity of 0.98/0.83, while the specific definition had a sensitivity/specificity of 0.77/0.99. Patients were classified into 3 groups: group 0 did not meet either definition, group 1 met the sensitive but not specific definition, and group 2 met both. The two-algorithm indicator approach adjusted using indicators for groups 1 and 2, while the two-algorithm restriction approach excluded patients in group 1 and adjusted using an indicator for group 2. Adjusted exposure odds ratios (ORs) were estimated for each approach using logistic regression.
Results:
The crude OR was 1.33 (95% CI, 1.07-1.63). Adjusting for the specific or sensitive definitions resulted in ORs of 1.09 (95% CI, 0.87-1.35) and 1.14 (95% CI, 0.91-1.40). The two-algorithm indicator method returned an OR of 1.07 (95% CI, 0.86-1.33). The two-algorithm restriction approach returned an OR of 1.02 (95% CI, 0.79-1.29) but excluded 20% of the cohort.
Conclusions:
The two-algorithm indicator approach may improve adjustment for claims-based confounders by returning a point estimate at least as unbiased as the better of the two component definitions.
Related Concept Videos
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Comparing the Survival Analysis of Two or More Groups
Study Design in Statistics
Does aspirin reduce the risk of heart attacks? Is one brand of fertilizer more effective at growing roses than another? Is fatigue as dangerous to a driver as the influence of alcohol? Questions like these are answered using randomized experiments with proper...
Assumptions of Survival Analysis

