Related Experiment Video
Updated: May 30, 2025

A Novel Method for Involving Women of Color at High Risk for Preterm Birth in Research Priority Setting
Published on: January 12, 2018
Fairness in Low Birthweight Predictive Models: Implications of Excluding Race/Ethnicity
Clare C Brown1, Michael Thomsen2, Benjamin C Amick3
1Department of Health Policy and Management, Fay W Boozman College of Public Health, University of Arkansas for Medical Sciences, 4301 W Markham St Slot #820-12, Little Rock, AR, 72205, USA. ccbrown@uams.edu.
Context:
To evaluate algorithmic fairness in low birthweight predictive models.
Study Design:
This study analyzed insurance claims (n = 9,990,990; 2013-2021) linked with birth certificates (n = 173,035; 2014-2021) from the Arkansas All Payers Claims Database (APCD).
Methods:
Low birthweight (< 2500 g) predictive models included four approaches (logistic, elastic net, linear discriminate analysis, and gradient boosting machines [GMB]) with and without racial/ethnic information. Model performance was assessed overall, among Hispanic individuals, and among non-Hispanic White, Black, Native Hawaiian/Other Pacific Islander, and Asian individuals using multiple measures of predictive performance (i.e., AUC [area under the receiver operating characteristic curve] scores, calibration, sensitivity, and specificity).
Results:
AUC scores were lower (underperformed) for Black and Asian individuals relative to White individuals. In the strongest performing model (i.e., GMB), the AUC scores for Black (0.718 [95% CI: 0.705-0.732]) and Asian (0.655 [95% CI: 0.582-0.728]) populations were lower than the AUC for White individuals (0.764 [95% CI: 0.754-0.775 ]). Model performance measured using AUC was comparable in models that included and excluded race/ethnicity; however, sensitivity (i.e., the percent of records correctly predicted as "low birthweight" among those who actually had low birthweight) was lower and calibration was weaker, suggesting underprediction for Black individuals when race/ethnicity were excluded.
Conclusions:
This study found that racially blind models resulted in underprediction and reduced algorithmic performance, measured using sensitivity and calibration, for Black populations. Such under prediction could unfairly decrease resource allocation needed to reduce perinatal health inequities. Population health management programs should carefully consider algorithmic fairness in predictive models and associated resource allocation decisions.
More Related Videos
Related Concept Videos
Regression Toward the Mean
Bias in Epidemiological Studies
z Scores and Area Under the Curve
Expected Frequencies in Goodness-of-Fit Tests
One-Way ANOVA: Unequal Sample Sizes
Punnett Squares

