Related Experiment Video
Updated: Jan 10, 2026

Quantified Assessment of Infant's Gross Motor Abilities Using a Multisensor Wearable
Published on: May 17, 2024
Identifying Predictors of Utilization of Skilled Birth Attendance in Uganda Through Interpretable Machine Learning
Shaheen M Z Memon1, Robert Wamala2, Ignace H Kabano1
1African Centre of Excellence in Data Science, College of Business and Economics, University of Rwanda, Kigali P.O. Box 4285, Rwanda.
None:
Skilled Birth Attendance (SBA) is essential for reducing maternal and neonatal mortality, yet access remains limited in many low- and middle-income countries. This study used machine learning to predict SBA use among Ugandan women and identify key influencing factors. We analyzed data from the 2016 Uganda Demographic and Health Survey, focusing on women aged 15 to 49 who had given birth in the preceding five years. After preparing and selecting relevant features, six tree-based models (decision tree, random forest, gradient boosting, XGBoost, LightGBM, CatBoost) and logistic regression were applied. Class imbalance was addressed using cost-sensitive learning, and hyperparameters were tuned via Bayesian optimization. XGBoost performed best (F1-score: 0.52; recall: 0.73; AUC: 0.75). SHapley Additive Explanations (SHAP) were used to interpret model predictions. Key predictors of SBA use included education level, antenatal care visits, region (especially Northern Uganda), perceived distance to a healthcare facility, and urban or rural residence. The results demonstrate the value of interpretable machine learning for identifying at-risk populations and guiding targeted maternal health interventions in Uganda.
Related Concept Videos
Steps in Outbreak Investigation
Regression Toward the Mean
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
