Related Experiment Video
Updated: Mar 12, 2026

An R-Based Landscape Validation of a Competing Risk Model
Published on: September 16, 2022
Four hundred or more participants needed for stable contingency table estimates of clinical prediction rule
Peter Kent1, Eleanor Boyle2, Jennifer L Keating3
1School of Physiotherapy and Exercise Science, Curtin University, Kent Street, Bently, Perth, Western Australia 6102, Australia; Clinical Biomechanics Research Unit, Department of Sports Science and Clinical Biomechanics, University of Southern Denmark, Campusvej 55, Odense M 5230, Denmark.
Objectives:
To quantify variability in the results of statistical analyses based on contingency tables and discuss the implications for the choice of sample size for studies that derive clinical prediction rules.
Study Design And Setting:
An analysis of three pre-existing sets of large cohort data (n = 4,062-8,674) was performed. In each data set, repeated random sampling of various sample sizes, from n = 100 up to n = 2,000, was performed 100 times at each sample size and the variability in estimates of sensitivity, specificity, positive and negative likelihood ratios, posttest probabilities, odds ratios, and risk/prevalence ratios for each sample size was calculated.
Results:
There were very wide, and statistically significant, differences in estimates derived from contingency tables from the same data set when calculated in sample sizes below 400 people, and typically, this variability stabilized in samples of 400-600 people. Although estimates of prevalence also varied significantly in samples below 600 people, that relationship only explains a small component of the variability in these statistical parameters.
Conclusion:
To reduce sample-specific variability, contingency tables should consist of 400 participants or more when used to derive clinical prediction rules or test their performance.
More Related Videos
Related Concept Videos
Contingency Table
Receiver Operating Characteristic Plot
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Hazard Ratio
For example, in a clinical trial...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Sensitivity, Specificity, and Predicted Value
Sensitivity is the...

