Related Experiment Video
Updated: Jan 9, 2026

Implementation of a Real-Time Psychosis Risk Detection and Alerting System Based on Electronic Health Records using CogStack
Published on: May 15, 2020
A microsimulation-based framework for mitigating societal bias in primary care data
Purpose:
The data generating mechanisms underlying health care data are infrequently considered, leading to inequitable equilibria being reinforced throughout the care continuum. As race-based criteria are reassessed, the effect of those criteria on patterns of disease progression should also be reevaluated. We proposed a novel microsimulation-based framework for attenuating societal bias in primary care registry data to study this.
Methods:
Our data transformation framework enables generating counterfactual outcome distributions that would have been observed in the absence of race-based diagnosis and treatment criteria. We developed a continuous-time, discrete-event individual-level simulation model of kidney function decline, measured by estimated glomerular filtration rate (eGFR). The model simulates individual eGFR trajectories over time. eGFR decline is accelerated by hypertension, diabetes, and reaching chronic kidney disease stage 3a, and can be delayed by interventions, which are applied based on eGFR level, measured with or without an adjustment for Black race. A Bayesian calibration procedure was applied to identify rates of eGFR decline corresponding to stage distributions in the cohort.
Results:
Under the counterfactual scenario without a race adjustment, Black individuals qualify for diagnosis earlier, and non-Black individuals later, than under the reference scenario with race adjustment. The difference was largest for earlier stages and smaller at each consecutive stage. We do not observe differences in life expectancy between the two scenarios.
Limitations:
Large variability in the prevalence of treatment and heterogeneity in treatment effectiveness may impact our results.
Conclusions:
Our data transformation framework demonstrates how the explicit representation of the data generation process could inform the effect of policy changes on clinical data distributions. The framework can flexibly be adapted to mitigate bias in other health data.
More Related Videos
06:55Inverse Probability of Treatment Weighting Propensity Score using the Military Health System Data Repository and National Death Index
Published on: January 8, 2020
11:21Methodology for Establishing a Community-Wide Life Laboratory for Capturing Unobtrusive and Continuous Remote Activity and Health Data
Published on: July 27, 2018
Related Concept Videos
Bias in Epidemiological Studies
Strategies for Assessing and Addressing Confounding
Confounding can be addressed at both the design phase of a study and through analytical methods after data...
Methods of Documentation VI: Case Management Model
For example, a patient with a chronic...
Secondary Healthcare System
Study Designs in Epidemiology
Observational studies are those where the researcher does not intervene but rather observes natural variations. They include cross-sectional, cohort, and...
Bias
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...