在帕特森的f统计中解释了姐妹排斥现象
Gözde Atağ1,2, Shamam Waldman3, Shai Carmi4
1Department of Biological Sciences, Middle East Technical University, Ankara 06800, Turkey.
Genetics
|September 18, 2024
概括
帕特森的f-统计数据可以显示"姐妹排斥",在那里,相关的群体似乎更接近远方的人群. 这项研究解释了这种模式是来自外部来源的基因流的结果,影响了人口统计学推断.
科学领域:
- 人口遗传学 人口遗传学
- 古代的基因组学 古代的基因组学
- 生物信息学是一种生物信息学.
背景情况:
- 帕特森的f统计 (f3和f4统计) 被广泛用于从全基因组等位基因频率数据进行人口统计推断.
- 这些统计数据也用于基于共享历史的人口聚类.
- 已经观察到一种被称为"姐妹排斥"的不寻常模式,在这种模式下,地理上接近的种群对更遥远的种群表现出更高的遗传亲和力,而不是彼此.
研究的目的:
- 在人口遗传分析中识别和解释"姐妹排斥"现象.
- 为了调查青铜时代东亚纳托利亚和希腊基因组中的姐妹排斥的一个新例.
- 提出和验证一个能够解释这种反直觉的遗传亲和关系模式的人口模型.
主要方法:
- 使用f3和f4统计数据对全基因组等位基因频率数据的分析.
- 主要成分分析 (PCA) 和多维缩放 (MDS) 在遗传距离上的应用.
- 模拟各种人口模型下的遗传数据,包括基因流动场景.
- 在模拟数据上计算f统计数据,以测试拟议的模型.
主要成果:
- 发现了一种新的姐妹排斥病例,青铜时代的东亚纳托利亚基因组与青铜时代的希腊相比,更有亲和力.
- 这种模式与考古/历史数据和PCA和MDS等替代遗传相似性测量方法的预期相矛盾.
- 模拟证实,来自遗传遥远来源的低水平基因流动或单向基因流动可以在f统计中产生姐妹排斥.
- 在MDS分析中,一致地将姐妹种群聚集在一起,突出了与f-统计结果的差异.
结论:
- 在使用f统计时,低级别的混合事件可以显著影响人口推理.
- 在f统计中观察到的"姐妹排斥"模式可能来自外部基因流,不一定反映真实的人口关系.
- 研究人员在解释f统计结果时应谨慎,尤其是在存在微妙的混合物时.
- 像MDS这样的替代方法可以在某些混合场景中更直观地表示遗传相似性.
相关概念视频
F Distribution
8.8K
The F distribution was named after Sir Ronald Fisher, an English statistician. The F statistic is a ratio (a fraction) with two sets of degrees of freedom; one for the numerator and one for the denominator. The F distribution is derived from the Student's t distribution. The values of the F distribution are squares of the corresponding values of the t distribution. One-Way ANOVA expands the t test for comparing more than two groups. The scope of that derivation is beyond the level of this...
8.8K
Bonferroni Test
2.6K
The Bonferroni test is a statistical test named after Carlo Emilio Bonferroni, an Italian mathematician best known for Bonferroni inequalities. This statistical test is a type of multiple comparison test to determine which means are different than the rest. Bonferroni test can minimize the Type 1 error by reducing the significance level alpha, which otherwise increases with sample pairs.
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
2.6K
Identifying Statistically Significant Differences: The F-Test
3.6K
The F-test is used to compare two sample variances to each other or compare the sample variance to the population variance. It is used to decide whether an indeterminate error can explain the difference in their values. The underlying assumptions that allow the use of the F-test include the data set or sets are normally distributed, and the data sets are independent of each other. The test statistic F is calculated by dividing one variance by another. In other words, the square of one standard...
3.6K
Behrens–Fisher Test
347
The Behrens-Fisher test is a statistical method designed to address the Behrens-Fisher problem, which arises when comparing the means of two normally distributed populations with unequal variances. Unlike the Student's t-test, which assumes equal variances, the Behrens-Fisher test allows for mean comparison without this restrictive assumption. This flexibility makes it particularly valuable in scenarios where two independent samples exhibit normality but lack variance homogeneity.
This test...
This test...
347
Fisher's Exact Test
1.4K
Fisher's exact test is a statistical significance test widely used to analyze 2x2 contingency tables, particularly in situations where sample sizes are small. Unlike the chi-squared test, which approximates P-values and assumes minimum expected frequencies of at least five in each cell, Fisher's exact test calculates the exact probability (P-value) of observing the data or more extreme results under the null hypothesis. This feature makes it especially valuable when the assumptions of...
1.4K
Friedman Two-way Analysis of Variance by Ranks
595
Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures...
595


