Control of population stratification by correlation-selected principal components.

Seunggeun Lee1, Fred A Wright, Fei Zou

  • 1Department of Biostatistics, University of North Carolina, Chapel Hill, North Carolina 27599, USA. slee@bios.unc.edu

Biometrics
|December 8, 2010
PubMed
Summary

Population stratification in genome-wide association studies inflates test statistics. A new EigenCorr method selects principal components based on phenotype correlation, improving power and saving computation.

Related Concept Videos

Stratified Sampling Method01:16

Stratified Sampling Method

Sampling is a technique to select a portion (or subset) of the larger population and study that portion (the sample) to gain information about the population. The sampling method ensures that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a stratified sample, divide the population into groups called strata and then take a...
Cluster Sampling Method01:20

Cluster Sampling Method

Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Friedman Two-way Analysis of Variance by Ranks01:21

Friedman Two-way Analysis of Variance by Ranks

Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures from...
Frequency-dependent Selection01:21

Frequency-dependent Selection

When the fitness of a trait is influenced by how common it is (i.e., its frequency) relative to different traits within a population, this is referred to as frequency-dependent selection. Frequency-dependent selection may occur between species or within a single species. This type of selection can either be positive—with more common phenotypes having higher fitness—or negative, with rarer phenotypes conferring increased fitness.Positive Frequency-Dependent SelectionIn positive...
Mechanistic Models: Compartment Models in Individual and Population Analysis01:23

Mechanistic Models: Compartment Models in Individual and Population Analysis

Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least squares (OLS)...
Three-Dimensional Analysis of Strain01:29

Three-Dimensional Analysis of Strain

Three-dimensional strain analysis is crucial for understanding how materials deform under stress, particularly in elastic, homogeneous materials. This method employs principal stress axes to simplify complex stress states into more understandable forms. Subjected to stress, a small cubic element within a material either expands or contracts along these axes, transforming into a rectangular parallelepiped. This transformation effectively illustrates the material's deformation. The principal...