减少基因组估计育种值在单步基因组最佳线性无偏预测器中估计可靠性的计算时间,使用不同的核心大小用于已验证和年轻的算法
S N Sanchez-Sierra1, Matias Bermann1, Natascha Vukasinovic2
1Department of Animal and Dairy Science, University of Georgia, Athens, GA 30602.
JDS communications
|March 6, 2026
概括
在APY核心集中减少基因型动物的数量,在单步基因最佳线性无偏预测 (ssGBLUP) 中显著加快了基因组估计育种价值 (GEBV) 可靠性计算. 这种优化加快了计算速度,但并没有大大降低可靠性近似的精度.
科学领域:
- 动物育种与遗传学
- 基因组预测 基因组预测
- 计算生物学 计算生物学
背景情况:
- 单步基因组最佳线性无偏预测 (ssGBLUP) 对于计算牲畜的基因组估计繁殖值 (GEBV) 是至关重要的.
- 由于矩阵反转要求,计算GEBV可靠性是计算密集的,特别是在大型基因型群体中.
- 经验证和年轻的算法 (APY) 通过利用稀疏矩阵结构,提供了一种方法来近似GEBV可靠性.
研究的目的:
- 为了减少在ssGBLUP.UP中近似计算GEBV可靠性的计算时间.
- 调查降低APY核心集大小对可靠性近似度精度的影响.
- 优化GEBV可靠性计算中的计算效率和准确性之间的平衡.
主要方法:
- 在牛犊呼吸道疾病的大型霍尔斯坦数据集上使用ssGBLUP与APY算法对GEBV的近似可靠性.
- 评估了与25k基准对比的不同APY核心集大小 (25k,20k,15k,10k,5k).
- 测量了不同核心大小和基准指标的近似可靠性之间的相关性和回归系数.
主要成果:
- 大致可靠性与基准之间的相关性在0.94至1.00之间,表明高度一致.
- 5k核心集实现了最快的计算时间 (55.02分钟),与25k基准 (171.27分钟) 相比减少了3.1倍.
- 减少核心集大小导致内存使用量减少2.1倍,在近似精度上有轻微的权衡.
结论:
- 减少APY核心集大小是一种有效的策略,可以加速ssGBLUP中的GEBV可靠性计算.
- 一个较小的核心集提供了显著的计算和内存节省,特别是在多特征分析.
- 在加速计算的同时,近似可靠性的调整突显了计算效率和预测准确性之间的固有权衡.
相关概念视频
Estimating Population Standard Deviation
3.4K
When the population standard deviation is unknown and the sample size is large, the sample standard deviation s is commonly used as a point estimate of σ. However, it can sometimes under or overestimate the population standard deviation. To overcome this drawback, confidence intervals are determined to estimate population parameters and eliminate any calculation bias accurately. However, this only applies to random samples from normally distributed populations. Knowing the sample mean and...
3.4K
Estimating Population Mean with Known Standard Deviation
9.8K
To construct a confidence interval for a single unknown population mean μ, where the population standard deviation is known, we need sample mean as an estimate for μ and we need the margin of error. Here, the margin of error (EBM) is called the error bound for a population mean (abbreviated EBM). The sample mean is the point estimate of the unknown population mean μ.
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
9.8K
Sample Size Calculation
6.8K
Knowledge of the sample size is the first requirement to conduct random sampling or an experiment. The sample size is the total number of units, observations, or groups (in some cases) used to get the data to estimate a population parameter. As the name suggests, the sample size is that of the sample drawn from the population and differs from the population size.
The sample size for the given experiment or sampling effort is fundamental to any study design. Sample size decides the number of...
The sample size for the given experiment or sampling effort is fundamental to any study design. Sample size decides the number of...
6.8K
Estimating Population Mean with Unknown Standard Deviation
9.0K
In practice, we rarely know the population standard deviation. In the past, when the sample size was large, this did not present a problem to statisticians. They used the sample standard deviation s as an estimate for σ and proceeded as before to calculate a confidence interval with close enough results. However, statisticians ran into problems when the sample size was small. A small sample size caused inaccuracies in the confidence interval.
William S. Gosset (1876–1937) of the...
William S. Gosset (1876–1937) of the...
9.0K
Improving Translational Accuracy
15.3K
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
15.3K
Improving Translational Accuracy
3.7K
3.7K


