Related Experiment Video
Updated: Jul 6, 2026

A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Prediction model of preeclampsia using machine learning based methods: a population based cohort study in China
Taishun Li1,2, Mingyang Xu3, Yuan Wang1
1Department of Obstetrics and Gynecology, Nanjing Drum Tower Hospital, The Affiliated Hospital of Nanjing University Medical School, Nanjing, China.
Insights
Machine learning models accurately predict preeclampsia using early pregnancy markers. The Voting Classifier showed superior performance for preterm preeclampsia, aiding early intervention and improving maternal and fetal outcomes.
Area of Science:
- Obstetrics and Gynecology
- Medical Informatics
- Maternal-Fetal Medicine
Background:
- Preeclampsia is a significant cause of maternal and perinatal morbidity with unknown pathogenesis.
- Early identification and intervention, such as aspirin therapy, are crucial for preeclampsia prevention.
- Developing effective prediction models is vital for early screening and risk stratification.
Purpose of the Study:
- To develop a robust machine learning-based prediction model for preeclampsia.
- To identify high-risk pregnancies for preeclampsia using early gestational markers.
- To provide an effective tool for early screening and prediction of preeclampsia.
Main Methods:
- A prospective cohort study included 5116 pregnant women.
- Maternal characteristics, biophysical, and biochemical markers (MAP, UtPI, PAPP-A, PLGF) were collected at 11-13+6 weeks' gestation.
- Five machine learning algorithms (Logistic Regression, Extra Trees, Voting, Gaussian Process, Stacking) were applied and validated using cross-validation.
Main Results:
- The Voting Classifier demonstrated superior performance in predicting preterm preeclampsia (AUC=0.884, DR=0.625 at 10% FPR).
- PLGF contribution was higher than PAPP-A in predicting overall preeclampsia, with the opposite trend for preterm preeclampsia.
- Machine learning model performance was comparable to the Fetal Medicine Foundation competing risk model.
Conclusions:
- Developed machine learning models offer an accessible tool for large-scale preeclampsia screening.
- Early prediction can help reduce the disease burden of preeclampsia.
- Improved maternal and fetal outcomes are anticipated through timely intervention based on accurate predictions.
Introduction:
Preeclampsia is a disease with an unknown pathogenesis and is one of the leading causes of maternal and perinatal morbidity. At present, early identification of high-risk groups for preeclampsia and timely intervention with aspirin is an effective preventive method against preeclampsia. This study aims to develop a robust and effective preeclampsia prediction model with good performance by machine learning algorithms based on maternal characteristics, biophysical and biochemical markers at 11-13 + 6 weeks' gestation, providing an effective tool for early screening and prediction of preeclampsia.
Methods:
This study included 5116 singleton pregnant women who underwent PE screening and fetal aneuploidy from a prospective cohort longitudinal study in China. Maternal characteristics (such as maternal age, height, pre-pregnancy weight), past medical history, mean arterial pressure, uterine artery pulsatility index, pregnancy-associated plasma protein A, and placental growth factor were collected as the covariates for the preeclampsia prediction model. Five classification algorithms including Logistic Regression, Extra Trees Classifier, Voting Classifier, Gaussian Process Classifier and Stacking Classifier were applied for the prediction model development. Five-fold cross-validation with an 8:2 train-test split was applied for model validation.
Results:
We ultimately included 49 cases of preterm preeclampsia and 161 cases of term preeclampsia from the 4644 pregnant women data in the final analysis. Compared with other prediction algorithms, the AUC and detection rate at 10% FPR of the Voting Classifier algorithm showed better performance in the prediction of preterm preeclampsia (AUC=0.884, DR at 10%FPR=0.625) under all covariates included. However, its performance was similar to that of other model algorithms in all PE and term PE prediction. In the prediction of all preeclampsia, the contribution of PLGF was higher than PAPP-A (11.9% VS 8.7%), while the situation was opposite in the prediction of preterm preeclampsia (7.2% VS 16.5%). The performance for preeclampsia or preterm preeclampsia using machine learning algorithms was similar to that achieved by the fetal medicine foundation competing risk model under the same predictive factors (AUCs of 0.797 and 0.856 for PE and preterm PE, respectively).
Conclusions:
Our models provide an accessible tool for large-scale population screening and prediction of preeclampsia, which helps reduce the disease burden and improve maternal and fetal outcomes.
Related Concept Videos
Mechanistic Models: Compartment Models in Individual and Population Analysis
Steps in Outbreak Investigation

