不同机器学习算法在预测产后护理利用率方面的性能评估和比较分析:来自2016年埃塞俄比亚人口和健康调查的证据
Daniel Niguse Mamo1, Agmasie Damtew Walle2, Eden Ketema Woldekidan1
1Department of Health Informatics, School of Public Health, College of Medicine and Health Sciences, Arbaminch University, Arbaminch, Ethiopia.
PLOS digital health
|January 9, 2025
概括
机器学习模型有效地预测了埃塞俄比亚的产后护理利用率. 地区和教育等关键因素显著地影响了获取,指导了针对母亲和新生儿的有针对性的健康倡议.
科学领域:
- 公共卫生 公共卫生
- 医疗保健中的机器学习
- 孕产妇和儿童的健康
背景情况:
- 产后护理 (PNC) 在产后前六周至关重要,这是产后和新生儿死亡率高的时期.
- 在被研究的国家中,近40%的妇女错过了产后检查,这凸显了护理利用率的差距.
- 预测PNC利用率对于开发有针对性的干预措施来改善孕产妇和新生儿的结果至关重要.
研究的目的:
- 评估和比较用于预测埃塞俄比亚产后护理利用的机器学习算法.
- 确定影响产后护理采用的主要人口和社会经济因素.
主要方法:
- 来自2016年埃塞俄比亚人口和健康调查 (7,193名女性) 的二次数据的分析.
- 应用十五个机器学习算法,用十倍交叉验证和合成少数群体过量采样技术来解决类不平衡.
- 具有特点的重要技术,以确定产后护理利用的重要预测因素.
主要成果:
- MLP分类器,随机森林分类器和包装分类器表现出色 (F1分数≥0.9498,AUC≥0.98).
- 确定的主要预测因素包括地区,居住地,孕产妇教育,宗教,财富指数,医疗保险和分娩地点.
- 十倍交叉验证与合成少数群体过量采样技术在处理类不平衡方面表现出优势.
结论:
- 机器学习模型可以准确预测产后护理使用情况,帮助公共卫生规划.
- 区域和社会经济因素是产后护理利用的关键决定因素.
- 解决阶级不平衡对于开发可靠的医疗预测模型至关重要.
相关概念视频
Regression Toward the Mean
6.3K
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when...
6.3K
Comparing the Survival Analysis of Two or More Groups
146
Survival analysis is a cornerstone of medical research, used to evaluate the time until an event of interest occurs, such as death, disease recurrence, or recovery. Unlike standard statistical methods, survival analysis is particularly adept at handling censored data—instances where the event has not occurred for some participants by the end of the study or remains unobserved. To address these unique challenges, specialized techniques like the Kaplan-Meier estimator, log-rank test, and...
146
Statistical Methods for Analyzing Epidemiological Data
299
Epidemiological data primarily involves information on specific populations' occurrence, distribution, and determinants of health and diseases. This data is crucial for understanding disease patterns and impacts, aiding public health decision-making and disease prevention strategies. The analysis of epidemiological data employs various statistical methods to interpret health-related data effectively. Here are some commonly used methods:
299
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
40
Mechanistic models play a crucial role in algorithms for numerical problem-solving, particularly in nonlinear mixed effects modeling (NMEM). These models aim to minimize specific objective functions by evaluating various parameter estimates, leading to the development of systematic algorithms. In some cases, linearization techniques approximate the model using linear equations.
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
40
Kaplan-Meier Approach
90
The Kaplan-Meier estimator is a non-parametric method used to estimate the survival function from time-to-event data. In medical research, it is frequently employed to measure the proportion of patients surviving for a certain period after treatment. This estimator is fundamental in analyzing time-to-event data, making it indispensable in clinical trials, epidemiological studies, and reliability engineering. By estimating survival probabilities, researchers can evaluate treatment effectiveness,...
90
Statistical Software for Data Analysis and Clinical Trials
491
Statistical software is pivotal in data analysis and clinical trials by providing tools to analyze data, draw conclusions, and make predictions. These software packages range from simple data management applications to complex analytical platforms, supporting various statistical tests, models, and simulation techniques. Their significance lies in their ability to handle vast amounts of data with precision and efficiency, enabling researchers to validate hypotheses, identify trends, and make...
491


