Related Experiment Video
Updated: May 29, 2025

Rare Event Detection Using Error-corrected DNA and RNA Sequencing
Published on: August 3, 2018
Optimal Control of Directional False Discovery Rates in Large-Scale Testing
Guozhu Tang1, Yicheng Kang2, Dongdong Xiang1
1KLATASDS-MOE, School of Statistics, East China Normal University, Shanghai, China.
This study introduces a novel three-group model for analyzing gene expression data, improving the identification of over-expressed and under-expressed genes while controlling false discoveries.
Area of Science:
- Biomedical data analysis
- Genomics
- Statistical genetics
Background:
- High-throughput technologies measure thousands of gene expression levels simultaneously.
- Identifying over-expressed and under-expressed genes is crucial for analyzing gene expression data.
- Existing two-group models fail to control specific false discovery rates for over- and under-expressed genes.
Purpose of the Study:
- To propose a general three-group model for gene expression data analysis.
- To develop a decision rule that controls both over- and under-expressed false discovery rates.
- To optimize the expected number of true discoveries while maintaining desired false discovery proportions.
Main Methods:
- Development of a general three-group model accommodating dependence between test statistics.
- Design of a decision rule with a monotonic structure for controlling false discovery rates.
- Linearization of two-directional false discovery rate constraints using the monotonic structure.
Main Results:
- The proposed decision rule optimizes true discoveries while controlling false discovery rates for both over- and under-expression.
- Data-driven versions of the procedures are suggested and their consistency is established.
- The new procedures demonstrate strong performance in comparisons with existing methods and in genomic applications.
Conclusions:
- The proposed three-group model and decision rule offer improved control over false discoveries in gene expression analysis.
- This approach enhances the reliability of identifying differentially expressed genes.
- The findings have significant implications for genomic studies and biomedical research.
More Related Videos
Related Concept Videos
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Errors In Hypothesis Tests
Detection of Gross Error: The Q Test
Testing a Claim about Mean: Unknown Population SD
Estimating a population mean requires the samples to be approximately normally distributed. The data should be collected from the randomly selected samples having no sampling bias. There is no specific requirement for sample size. But if the sample size is less than 30, and we don't know the population standard deviation, a different approach is used;...

