后期预测测试的预期行为及其意想不到的解释
Luiza Guimarães Fabreti1,2, Lyndon M Coghill3,4, Robert C Thomson5
1GeoBio-Center, Ludwig-Maximilians-Universität München, Richard-Wagner-Str. 10, Munich 80333, Germany.
Molecular biology and evolution
|March 4, 2024
概括
贝叶斯后期预测对于检测进化研究中模型不适合的情况至关重要. 本研究描述了其预期的P值分布,揭示了非均性,并强调了其用于评估进化模型的实用性.
科学领域:
- 进化生物学 进化生物学
- 计算生物学 计算生物学
- 统计建模 统计建模
背景情况:
- 进化模型和经验数据之间的不匹配可能导致偏见的结论.
- 贝叶斯后置预测提供了一种灵活的方法来检测模型不适合的情况.
- 进化建模中后期预测测试的预期行为仍然没有表征.
研究的目的:
- 描述进化模型后期预测P值的预期分布.
- 澄清进化数据分析中后期预测测试的解释.
主要方法:
- 对贝叶斯后置预测分布的分析.
- 在模型适配条件下P值分布的表征.
- 与频率的P值属性进行比较.
主要成果:
- 后预测P值的预期分布通常不均.
- 极端后期预测P值提供了对模型适合性差的重要证据.
- 拟合模型导致预期分布集中在中间值周围.
结论:
- 不均的P值分布并不妨碍后期预测测试的应用.
- 后预测P值代表了观察到的数据与观察到的数据一样极端的概率.
- 这项工作对于正确解释和应用贝叶斯模型在进化研究中的充分性至关重要.
相关概念视频
Sensitivity, Specificity, and Predicted Value
329
In healthcare diagnostics, laboratory tests play a crucial role in identifying and diagnosing a wide range of medical conditions. However, interpreting test results is not always straightforward. An abnormal test result does not always confirm the presence of a disease, just as a normal result does not guarantee its absence. To assess the reliability of these diagnostic tools, healthcare practitioners rely on two key statistical indicators: sensitivity and specificity.
Sensitivity is the...
Sensitivity is the...
329
Hindsight Biases
3.4K
Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now?
3.4K
Regression Toward the Mean
6.3K
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when...
6.3K
Unusual Results
3.2K
Unusual results are those that have a very low chance of occurring. Unusual results can be identified using probabilities and the range rule of thumb. In problems involving probability, unusual results can be observed in 2 instances – an unusually high number of successes or an unusually low number of successes.
According to the range rule of thumb, any value above or below two standard deviations, 2σ from the mean, μ is considered unusual.
Maximum unusual value =...
According to the range rule of thumb, any value above or below two standard deviations, 2σ from the mean, μ is considered unusual.
Maximum unusual value =...
3.2K
Testing a Claim about Population Proportion
3.3K
A complete procedure for testing a claim about a population proportion is provided here.
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
3.3K
Expected Frequencies in Goodness-of-Fit Tests
2.5K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
2.5K


