通过使用可行的通用最小平方来改进SNP遗传性的功能丰富估计
Zewei Xiong1, Thuan-Quoc Thach1, Yan Dora Zhang2
1Department of Psychiatry, Li Ka Shing Faculty of Medicine, The University of Hong Kong, Hong Kong SAR, China.
HGG advances
|February 8, 2024
概括
我们开发了通用链接不平衡得分回归 (g-LDSC) 以使用全基因组关联研究 (GWAS) 总结数据进行更精确的功能丰富估计. 这种方法通过利用整个LD矩阵来改进分层LDSC,从而产生更现实的结果.
科学领域:
- 遗传学 遗传学 是一个
- 统计遗传学 统计遗传学
- 生物信息学是一种生物信息学.
背景情况:
- 功能性丰富分析将生物途径与疾病病原和治疗点联系起来.
- 分层链接不平衡得分回归 (s-LDSC) 通过全基因组关联研究 (GWAS) 总结数据估计功能丰富,但仅部分使用链接不平衡 (LD) 信息.
研究的目的:
- 引入通用链接不平衡得分回归 (g-LDSC),一种用于估计功能丰富的新方法.
- 通过利用完整的LD矩阵来提高功能丰富估计的精度和现实性.
主要方法:
- g-LDSC利用了整个LD矩阵,并使用可行的概括最小平方估计来考虑相关的错误结构.
- 该方法与s-LDSC.相同的假设和回归模型制定.
- 在各种场景下进行模拟,将g-LDSC与s-LDSC进行比较.
主要成果:
- g-LDSC提供了比s-LDSC更精确的功能丰富估计,即使有模型错误规范.
- 对15个英国生物库特征的应用显示,与s-LDSC相比,g-LDSC的功能丰富估计较低,更为现实.
- 在15个特征中,g-LDSC发现的功能注释 (118) 比s-LDSC (51) 丰富得多.
结论:
- g-LDSC为基因研究中的功能丰富分析提供了更准确和更强大的方法.
- 该方法改善了与复杂特征相关的生物相关途径的识别.
- g-LDSC增强了GWAS总结数据的实用性,用于发现疾病机制和潜在的治疗点.
更多相关视频
相关概念视频
Genome-wide Association Studies-GWAS
13.4K
Genome-wide association studies or GWAS are used to identify whether common SNPs are associated with certain diseases. Suppose specific SNPs are more frequently observed in individuals with a particular disease than those without the disease. In that case, those SNPs are said to be associated with the disease. Chi-square analysis is performed to check the probability of the allele likely to be associated with the disease.
GWAS does not require the identification of the target gene involved in...
GWAS does not require the identification of the target gene involved in...
13.4K
Residuals and Least-Squares Property
7.4K
The vertical distance between the actual value of y and the estimated value of y. In other words, it measures the vertical distance between the actual data point and the predicted point on the line
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
7.4K
Comparing Copy Number Variations and SNPs
17.7K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
17.7K
Heritability
200
Heritability is a statistical concept that measures the degree to which genetic differences among individuals contribute to trait variations within a population. It is a fundamental idea in genetics, often prone to misinterpretation. Heritability is expressed as a percentage, reflecting the proportion of variation in a specific trait across a population that can be linked to genetic differences. However, it's important to understand that heritability does not determine how "genetic"...
200
Mechanistic Models: Compartment Models in Individual and Population Analysis
43
Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
43
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
507
This lesson introduces two critical methods in pharmacokinetics, the Wagner-Nelson and Loo-Riegelman methods, used for estimating the absorption rate constant (ka) for drugs administered via non-intravenous routes. The Wagner-Nelson method relates ka to the plasma concentration derived from the slope of a semilog percent unabsorbed time plot. However, it is limited to drugs with one-compartment kinetics and can be impacted by factors like gastrointestinal motility or enzymatic degradation.
On...
On...
507


