Penalized Logistic Regression Analysis for Genetic Association Studies of Binary Phenotypes

Ying Yu1, Siyuan Chen2, Samantha Jean Jones3

  • 1Department of Statistics and Actuarial Science, Simon Fraser University, Burnaby, British Columbia, Canada, ying_yu_5@sfu.ca.

Human Heredity
|June 29, 2022
PubMed
Summary

This study introduces a penalized logistic regression method using log-F priors to address data sparsity in genetic association studies. The new approach offers reduced bias and mean squared error for analyzing rare variants.

Related Concept Videos

Genome-wide Association Studies-GWAS01:11

Genome-wide Association Studies-GWAS

Genome-wide association studies or GWAS are used to identify whether common SNPs are associated with certain diseases. Suppose specific SNPs are more frequently observed in individuals with a particular disease than those without the disease. In that case, those SNPs are said to be associated with the disease. Chi-square analysis is performed to check the probability of the allele likely to be associated with the disease.
GWAS does not require the identification of the target gene involved in...
14.1K
Epistasis Analysis01:09

Epistasis Analysis

Although Mendel chose seven unrelated traits in peas to study gene segregation, most traits involve multiple gene interactions that create a spectrum of phenotypes. When the interaction of various genes or alleles at different locations influences a phenotype, this is called epistasis. Epistasis often involves one gene masking or interfering with the expression of another (antagonistic epistasis). Epistasis often occurs when different genes are part of the same biochemical pathway. The...
5.2K
Probability Laws01:49

Probability Laws

Overview
41.6K
Pedigree Analysis01:35

Pedigree Analysis

Overview
85.0K
Punnett Squares01:00

Punnett Squares

Overview
116.2K
Mechanistic Models: Compartment Models in Individual and Population Analysis01:23

Mechanistic Models: Compartment Models in Individual and Population Analysis

Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
85