Caution regarding the choice of standard deviations to guide sample size calculations in clinical trials

Henian Chen1, Nanhua Zhang, Xiaosun Lu

  • 1Department of Epidemiology & Biostatistics, College of Public Health, University of South Florida, Tampa, FL 33612, USA. hchen1@health.usf.edu

Abstract

Related Concept Videos

Testing a Claim about Standard Deviation01:19

Testing a Claim about Standard Deviation

A complete procedure to test a claim about population standard deviation or population variance is explained here.
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Chebyshev's Theorem to Interpret Standard Deviation01:15

Chebyshev's Theorem to Interpret Standard Deviation

Chebyshev’s theorem, also known as Chebyshev’s Inequality, states that the proportion of values of a dataset for K standard deviation is calculated using the equation:
Estimating Population Standard Deviation01:26

Estimating Population Standard Deviation

When the population standard deviation is unknown and the sample size is large, the sample standard deviation s is commonly used as a point estimate of σ. However, it can sometimes under or overestimate the population standard deviation. To overcome this drawback, confidence intervals are determined to estimate population parameters and eliminate any calculation bias accurately. However, this only applies to random samples from normally distributed populations. Knowing the sample mean and...
Standard Deviation of Calculated Results01:14

Standard Deviation of Calculated Results

Standard deviation measures the spread of data around the mean value. Many large data sets follow a Gaussian distribution, also known as a normal distribution. This distribution is bell-shaped curved, with the most frequently observed value (mean or central value) in the middle. The farther away from the central value, the greater the deviation from the central value, and the lower the frequency.
A broad Gaussian distribution curve has a wider standard deviation, representing a data set with...
Calculating Standard Deviation01:08

Calculating Standard Deviation

The standard deviation is the most common measure of variation. It is a value that tells us how far a data value is from the mean value in a dataset. Further, the standard deviation is always a positive value or zero.
The standard deviation value is small when all the data is concentrated close to the mean. Here the data exhibits low variation. The standard deviation value is larger when the data values are more spread out from the mean. Here, the data displays high variation.       
Let us...
Range Rule of Thumb to Interpret Standard Deviation01:13

Range Rule of Thumb to Interpret Standard Deviation

The range rule of thumb in statistics helps us calculate a dataset's minimum and maximum values with known standard deviation. This rule is based on the concept that 95% of all values in a dataset lie within two standard deviations from the mean.
For instance, the range rule of thumb can be used to find the tallest and the shortest student in a class, given the mean student height and standard deviation. If the mean student height is 1.6 m and the standard deviation, s is 0.05 m, the height of...