Related Experiment Video
Updated: May 7, 2026

Development of an Individual-Tree Basal Area Increment Model using a Linear Mixed-Effects Approach
Published on: July 3, 2020
Statistical inference for high-dimensional generalized estimating equations
1Department of Statistics and Probability, Michigan State University, 619 Red Cedar Road, East Lansing, MI, 48824, United States.
This study introduces a new statistical method for analyzing complex, high-dimensional correlated data, particularly useful in omics research. The procedure provides reliable confidence intervals for associations, improving insights from large datasets like those in COVID-19 proteomics.
Area of Science:
- Biostatistics
- Genomics
- Proteomics
Background:
- Analyzing correlated data with many variables is challenging, especially with limited sample sizes.
- High-throughput omics data, like proteomics, often present this high-dimensional correlated structure.
- Existing methods struggle with inference for high-dimensional regression coefficients in generalized estimating equations.
Purpose of the Study:
- To develop a novel inference procedure for linear functionals of high-dimensional regression coefficients.
- To address the analysis of correlated omics data, motivated by COVID-19 studies.
- To introduce a data-driven method for selecting tuning parameters in high-dimensional settings.
Main Methods:
- Developed a novel inference procedure using projected estimating equations.
- Established asymptotic normality of the proposed estimator under mild conditions.
- Introduced a cross-validation technique for tuning parameter selection.
Main Results:
- The proposed estimator is asymptotically normally distributed.
- Demonstrated robust finite-sample performance in simulations, particularly in bias and coverage.
- Successfully applied the procedure to provide confidence intervals for protein-COVID risk associations.
Conclusions:
- The novel procedure offers a statistically sound approach for high-dimensional correlated data analysis.
- The method enhances confidence interval estimation for associations in omics studies.
- The data-driven cross-validation improves practical application and reliability.
More Related Videos
09:23Quantification of Information Encoded by Gene Expression Levels During Lifespan Modulation Under Broad-range Dietary Restriction in C. elegans
Published on: August 16, 2017
08:27Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits
Published on: September 27, 2019
Related Concept Videos
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Statistical Methods for Analyzing Epidemiological Data
What are Estimates?
The estimate for the mean of a sample is denoted by ͞x, whereas the mean of the population is designated as μ. Further, parameters such...
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
On...
Distributions to Estimate Population Parameter
Parametric Survival Analysis: Weibull and Exponential Methods
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...