Related Experiment Video
Updated: Mar 8, 2026

A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Improved standard error estimator for maintaining the validity of inference in cluster randomized trials with a small
Whitney P Ford1, Philip M Westgate1
1Department of Biostatistics, College of Public Health, University of Kentucky, Lexington, KY, 40536, USA.
Abstract:
Cluster randomized trials (CRTs) are studies in which clusters of subjects are randomized to different trial arms. Due to the nature of outcomes within the same cluster to be correlated, generalized estimating equations (GEE) are growing as a popular choice for the analysis of data arising from CRTs. In the past, research has shown that analyses using GEE could result in liberal inference due to the use of the empirical sandwich covariance matrix estimator, which can yield negatively biased standard error estimates when the number of clusters is not large. Many techniques have been presented to correct this negative bias; however, use of these corrections can still result in biased standard error estimates and thus test sizes that are not consistently at their nominal level. Therefore, there is a need for an improved correction such that nominal type I error rates will consistently result. In this manuscript, we study the use of recently developed corrections for empirical standard error estimation and the use of a combination of two popular corrections. In an extensive simulation study, we found that nominal type I error rates can be consistently attained when using an average of two popular corrections developed by Mancl and DeRouen (, Biometrics 57, 126-134) and Kauermann and Carroll (, Journal of the American Statistical Association 96, 1387-1396). Therefore, use of this new correction was found to notably outperform the use of previously recommended corrections.
Related Concept Videos
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Estimating Population Mean with Unknown Standard Deviation
William S. Gosset (1876–1937) of the...
Estimating Population Mean with Known Standard Deviation
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
Estimating Population Standard Deviation
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...

