ScITree:从流行病学和基因组数据的传播树的可扩展贝叶斯推断
Hannah Waddel1, Katia Koelle2, Max S Y Lau1
1Department of Biostatistics and Bioinformatics, Emory University, Atlanta, Georgia, United States of America.
PLoS computational biology
|June 10, 2025
概括
我们开发了一种新的贝叶斯系系动力学模型,ScITree,它准确地推断出疾病传播动态. 这种可扩展的方法克服了计算瓶,使得大型疫情的有效分析能够及时响应公共卫生.
科学领域:
- 流行病学 流行病学
- 进化生物学 进化生物学
- 计算生物学 计算生物学
背景情况:
- 植物动力学模型整合了流行病学和进化数据来研究疾病爆发.
- 当前的方法通常使用近似,限制准确性和可扩展性.
- 之前的机械模型 (劳法) 准确但计算密集.
研究的目的:
- 开发一个可扩展和准确的贝叶斯系动力学模型,用于推断联合流行病学和进化动态.
- 克服现有的机械力学动力学方法的计算局限性.
- 为疫情分析提供一个易于部署的框架.
主要方法:
- 提出了一个可扩展的时空动力学框架 (ScITree),使用对突变的无限位点假设.
- 实施完整的贝叶斯方法,具有完全整合流行病学和进化过程的可能性.
- 开发了一个计算效率高的数据增量马尔科夫链蒙特卡洛算法用于参数推理.
主要成果:
- ScITree实现了与Lau方法相比较的高推断准确度.
- ScITree展示了显著的可扩展性,计算时间随着爆发规模的线性增加.
- 应用到口疫疫情爆发时,ScITree产生了可靠的传播动态估计.
结论:
- ScITree提供了一种计算效率高且高度可扩展的用于植物动力学分析的框架.
- 该模型有助于准确推断疾病传播的时空动态.
- 这种方法支持及时有效地应对疫情爆发.
相关概念视频
Causality in Epidemiology
331
Causality or causation is a fundamental concept in epidemiology, vital for understanding the relationships between various factors and health outcomes. Despite its importance, there's no single, universally accepted definition of causality within the discipline. Drawing from a systematic review, causality in epidemiology encompasses several definitions, including production, necessary and sufficient, sufficient-component, counterfactual, and probabilistic models. Each has its strengths and...
331
Statistical Software for Data Analysis and Clinical Trials
505
Statistical software is pivotal in data analysis and clinical trials by providing tools to analyze data, draw conclusions, and make predictions. These software packages range from simple data management applications to complex analytical platforms, supporting various statistical tests, models, and simulation techniques. Their significance lies in their ability to handle vast amounts of data with precision and efficiency, enabling researchers to validate hypotheses, identify trends, and make...
505
Statistical Methods for Analyzing Epidemiological Data
313
Epidemiological data primarily involves information on specific populations' occurrence, distribution, and determinants of health and diseases. This data is crucial for understanding disease patterns and impacts, aiding public health decision-making and disease prevention strategies. The analysis of epidemiological data employs various statistical methods to interpret health-related data effectively. Here are some commonly used methods:
313
Steps in Outbreak Investigation
108
In the ever-evolving field of public health, statistical analysis serves as a cornerstone for understanding and managing disease outbreaks. By leveraging various statistical tools, health professionals can predict potential outbreaks, analyze ongoing situations, and devise effective responses to mitigate impact. For that to happen, there are a few possible stages of the analysis:
108
Evolutionary Relationships through Genome Comparisons
5.7K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.7K
Mutation, Gene Flow, and Genetic Drift
58.2K
In a population that is not at Hardy-Weinberg equilibrium, the frequency of alleles changes over time. Therefore, any deviations from the five conditions of Hardy-Weinberg equilibrium can alter the genetic variation of a given population. Conditions that change the genetic variability of a population include mutations, natural selection, non-random mating, gene flow, and genetic drift (small population size).
58.2K


