Related Experiment Video
Updated: May 8, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
The effect of misclassification on sample size for one and two-sample tests with binary endpoints
Péter Hársfalvi1,2, Jenő Reiczigel1
1Department of Biostatistics, University of Veterinary Medicine Budapest, Budapest, Hungary.
Ignoring binary data misclassification in study design reduces statistical power. This study provides sample size formulas and R functions to adjust for misclassification (sensitivity and specificity) during study planning, ensuring adequate power.
Area of Science:
- Biostatistics
- Statistical Methods
- Epidemiology
Background:
- Analysis of binary data increasingly incorporates misclassification methods.
- Study designs often neglect potential misclassification due to a lack of sample size formulas and software.
- Ignoring misclassification during design can lead to significant power loss when addressed only during analysis.
Purpose of the Study:
- To emphasize the necessity of adjusting sample size for misclassification in the design phase of studies analyzing binary data.
- To provide a practical sample size calculation procedure for studies with binary endpoints, accounting for misclassification.
- To illustrate the impact of misclassification on required sample sizes for one-sample and two-sample tests.
Main Methods:
- Development of sample size formulas for one-sample and two-sample tests for binary endpoints, incorporating misclassification.
- Implementation of the sample size procedure as an R function.
- Calculation of sample sizes based on presumed binomial parameters, desired power, sensitivity (Se), and specificity (Sp).
Main Results:
- Misclassification significantly impacts the required sample size in both one-sample and two-sample testing scenarios.
- The developed R function provides a tool for researchers to calculate appropriate sample sizes.
- Comparison of sample sizes with and without misclassification highlights the potential for power loss.
Conclusions:
- Integrating misclassification correction into the study design phase, through appropriate sample size adjustment, is crucial.
- The provided methodology and R function can help researchers avoid power loss and design more robust studies.
- Accurate estimation of sensitivity and specificity is vital for effective sample size calculation in the presence of misclassification.
More Related Videos
Related Concept Videos
Errors In Hypothesis Tests
One-Way ANOVA: Unequal Sample Sizes
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Sample Size Calculation
The sample size for the given experiment or sampling effort is fundamental to any study design. Sample size decides the number of...
McNemar's Test

