Structural equation modelling (SEM) for malaria prevalence and risk factors in Uganda
Grace Kakaire1,2, Edna Chepkemoi Chumoh3, Mabula Salyungu3,4
1School of Statistics and Planning, Makerere University, Kampala, Uganda. kakairegrace2@gmail.com.
Malaria Journal
|November 7, 2025
Summary
Structural Equation Modelling revealed child health and environmental factors significantly impact malaria prevalence in children under five. Integrated interventions addressing child health and environmental exposures are crucial for reducing malaria burden in sub-Saharan Africa.
Area of Science:
- Epidemiology
- Public Health
- Biostatistics
Background:
- Malaria significantly affects children under five in sub-Saharan Africa, with complex contributing factors.
- Traditional analyses struggle to capture the multifaceted relationships influencing malaria transmission.
- Structural Equation Modelling (SEM) offers a robust approach to understanding these intricate pathways.
Purpose of the Study:
- To explore latent and observed predictors of child malaria prevalence using SEM.
- To provide a comprehensive understanding of the underlying factors influencing child malaria.
- To identify key pathways affecting malaria transmission in young children.
Main Methods:
- Utilized secondary data from the 2018-2019 Uganda Malaria Indicator Survey (MIS).
- Constructed a SEM framework with latent variables: Socioeconomic Status (SES), Environment, Maternal Health, and Child Health.
- Assessed model fit using standard indices (CFI, TLI, RMSEA, SRMR).
Main Results:
- Child Health showed the strongest positive association with malaria status (β=0.22, p<0.001).
- Environmental factors exhibited a marginal negative association (β=-0.36, p=0.056).
- Socioeconomic Status and Maternal Health were not statistically significant predictors.
Conclusions:
- SEM effectively disentangles complex factors influencing child malaria.
- Child health status and environmental factors are critical in malaria prevalence.
- Findings support integrated interventions targeting child health and environmental exposures to reduce malaria burden.
Related Concept Videos
Mechanistic Models: Compartment Models in Individual and Population Analysis
235
Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
235
Statistical Methods for Analyzing Epidemiological Data
888
Epidemiological data primarily involves information on specific populations' occurrence, distribution, and determinants of health and diseases. This data is crucial for understanding disease patterns and impacts, aiding public health decision-making and disease prevention strategies. The analysis of epidemiological data employs various statistical methods to interpret health-related data effectively. Here are some commonly used methods:
888
Steps in Outbreak Investigation
476
In the ever-evolving field of public health, statistical analysis serves as a cornerstone for understanding and managing disease outbreaks. By leveraging various statistical tools, health professionals can predict potential outbreaks, analyze ongoing situations, and devise effective responses to mitigate impact. For that to happen, there are a few possible stages of the analysis:
476
Causality in Epidemiology
1.5K
Causality or causation is a fundamental concept in epidemiology, vital for understanding the relationships between various factors and health outcomes. Despite its importance, there's no single, universally accepted definition of causality within the discipline. Drawing from a systematic review, causality in epidemiology encompasses several definitions, including production, necessary and sufficient, sufficient-component, counterfactual, and probabilistic models. Each has its strengths and...
1.5K
Multiple Regression
3.7K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
3.7K
Regression Analysis
7.9K
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
7.9K


