预测感染性腹的发生率与症状监测数据,使用基于堆叠的组合模型
Pengyu Wang1, Wangjian Zhang1, Hui Wang2
1Department of Medical Statistics, School of Public Health & Center for Health Information Research & Sun Yat-sen Global Health Institute, Sun Yat-sen University, Guangzhou, China.
BMC infectious diseases
|February 26, 2024
概括
预测传染性腹发病率对于公共卫生至关重要. 症状监测数据显著改善了传染性腹预测模型,优于气象数据,并通过堆叠组合方法带来更好的整体模型性能.
科学领域:
- 流行病学 流行病学
- 公共卫生 公共卫生
- 机器学习 机器学习
背景情况:
- 传染性腹是全球重要的公共卫生问题.
- 准确预测传染性腹发病率对于有效的公共卫生干预至关重要.
- 这项研究旨在利用先进的建模技术提高传染性腹的预测.
研究的目的:
- 开发和评估传染性腹发病率的预测模型.
- 为了比较症状监测数据与气象数据在提高预测准确性的有效性.
- 与单个基底模型相比,评估堆叠组合模型的性能.
主要方法:
- 利用广州 (2016-2021) 传染性腹病例,症状和气象因素的监测数据.
- 开发了四种基本预测模型:人工神经网络 (ANN),长期短期记忆 (LSTM) 网络,支持向量回归 (SVR) 和极端梯度增强 (XGBoost).
- 使用堆叠组合基础模型,创建最终的预测模型,用MAPE,RMSE和MAE进行评估.
主要成果:
- 与仅使用气象数据的模型相比,结合症状监测数据的模型显示出更高的性能 (较低的RMSE,MAE,MAPE).
- 在基础模型中,LSTM模型表现出了最佳的个人性能.
- 堆叠组合模型实现了最高的精度,RMSE,MAE和MAPE分别为75.82,55.93和15.70%,优于所有基本模型.
结论:
- 症状监测数据比气象数据更有效,可以提高传染性腹预测模型的准确性.
- 堆叠组合方法提供了一个强大的方法来结合多个模型,克服选择一个单一的最佳模型的挑战,并产生卓越的预测性能.
更多相关视频
06:55A High-throughput Platform for the Screening of Salmonella spp./Shigella spp.
Published on: November 7, 2018
9.0K
12:21A Mouse Model for the Transition of Streptococcus pneumoniae from Colonizer to Pathogen upon Viral Co-Infection Recapitulates Age-Exacerbated Illness
Published on: September 28, 2022
2.5K
相关概念视频
Steps in Outbreak Investigation
128
In the ever-evolving field of public health, statistical analysis serves as a cornerstone for understanding and managing disease outbreaks. By leveraging various statistical tools, health professionals can predict potential outbreaks, analyze ongoing situations, and devise effective responses to mitigate impact. For that to happen, there are a few possible stages of the analysis:
128
Statistical Software for Data Analysis and Clinical Trials
550
Statistical software is pivotal in data analysis and clinical trials by providing tools to analyze data, draw conclusions, and make predictions. These software packages range from simple data management applications to complex analytical platforms, supporting various statistical tests, models, and simulation techniques. Their significance lies in their ability to handle vast amounts of data with precision and efficiency, enabling researchers to validate hypotheses, identify trends, and make...
550
