Related Experiment Videos
Boosting the Power of the Sequence Kernel Association Test by Properly Estimating Its Null Distribution
1Department of Biostatistics, College of Public Health, University of Iowa, Iowa City, IA 52242, USA.
American Journal of Human Genetics
|June 14, 2016
Summary
A new method, SKAT+, improves rare-variant association studies by estimating null distribution using only control subjects. This enhances statistical power while maintaining type I error control compared to the standard Sequence Kernel Association Test (SKAT).
Area of Science:
- Genetics
- Statistical genetics
- Bioinformatics
Background:
- Rare-variant association studies are crucial for identifying genetic variants linked to diseases.
- The Sequence Kernel Association Test (SKAT) is a widely used statistical method in this field.
- Current SKAT methods face limitations in statistical power due to using all subjects for parameter estimation.
Purpose of the Study:
- To develop an improved estimation method for the null distribution in rare-variant association studies.
- To enhance the statistical power of the Sequence Kernel Association Test (SKAT) while preserving type I error control.
- To introduce SKAT+, a novel method utilizing control subjects for estimation.
Main Methods:
- Developed SKAT+ (Sequence Kernel Association Test plus), a novel estimation method for null distribution.
- Employed a test statistic identical to SKAT but with a modified estimation approach using only control subjects.
- Validated SKAT+ through extensive simulation studies and real-world data applications (Genetic Analysis Workshop 17, Ocular Hypertension Treatment Study).
Main Results:
- SKAT+ demonstrated superior statistical power compared to the standard SKAT method.
- The method effectively maintained control over the type I error rate.
- Applications to real genetic datasets confirmed the enhanced performance of SKAT+.
Conclusions:
- SKAT+ offers a more powerful approach for rare-variant association studies than traditional SKAT.
- The method's ability to use control subjects for estimation is key to its improved performance.
- SKAT+ is applicable to various extensions of SKAT in genetic research.
Related Concept Videos
Expected Frequencies in Goodness-of-Fit Tests
6.4K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
6.4K
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
373
Statistical inference techniques, paramount in hypothesis testing, differentiate into two broad categories: parametric and nonparametric statistics.
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
373
Distributions to Estimate Population Parameter
5.0K
The accurate values of population parameters such as population proportion, population mean, and population standard deviation (or variance) are usually unknown. These are fixed values that can only be estimated from the data collected from the samples. The estimates of each of these parameters are sample proportion, the sample mean, and sample standard deviation (or variance). To obtain the values of these sample statistics, data are required that have particular distribution and central...
5.0K
Significance Testing: Overview
10.6K
Significance testing is a set of statistical methods used to test whether a claim about a parameter is valid. In analytical chemistry, significance testing is used primarily to determine whether the difference between two values comes from determinate or random errors. The effect of a particular change in the measurement protocol, analyst, or sample itself can cause a deviation from the expected result. In the case of a suspected deviation/outlier, we need to be able to confirm mathematically...
10.6K
Wald-Wolfowitz Runs Test II
473
The Wald-Wolfowitz runs test, commonly referred to as the runs test, is a nonparametric test used to assess the randomness of ordered data. The test evaluates the number of runs, which are consecutive sequences of similar elements within the data. If the number of runs is significantly higher or lower than expected, the data is considered non-random, indicating a detectable pattern or structure.
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and 0s. In...
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and 0s. In...
473
Statistical Hypothesis Testing
5.6K
Hypothesis testing is a critical statistical procedure facilitating informed, evidence-based decisions. It begins with a hypothesis, which is a tentative explanation, or a prediction about a population parameter. This hypothesis can be either a null hypothesis (H0), indicating no effect or difference, or an alternative hypothesis (Ha), suggesting an effect or difference.
Statistical significance measures the probability that an observed result occurred by chance. If this probability, known as...
Statistical significance measures the probability that an observed result occurred by chance. If this probability, known as...
5.6K