Related Experiment Video
Updated: Sep 26, 2025

Novel Object Recognition and Object Location Behavioral Testing in Mice on a Budget
Published on: November 20, 2018
Current controversies: Null hypothesis significance testing
Philip M Sedgwick1, Anne Hammer2,3, Ulrik Schiøler Kesmodel4,5
1Institute for Medical and Biomedical Education, St George's, University of London, London, UK.
Null hypothesis significance testing (NHST) is widely used in healthcare but often misinterpreted. This commentary highlights NHST limitations and advocates for statistical reform in clinical research.
Area of Science:
- Medical Statistics
- Clinical Research Methodology
- Obstetrics and Gynecology
Background:
- Null hypothesis significance testing (NHST) with a 0.05 significance level is standard in healthcare, particularly in obstetric and gynecological research.
- NHST's application for inferring clinical significance from statistical significance is a common misunderstanding of its original purpose.
- Limitations of NHST, including sensitivity to sample size and susceptibility to Type I and II errors, are frequently overlooked.
Purpose of the Study:
- To critically examine the historical context and inherent limitations of traditional null hypothesis significance testing (NHST).
- To address the controversy surrounding the interpretation and application of NHST in clinical decision-making, especially in obstetric and gynecological research.
- To propose a reconsideration of current statistical practices and advocate for a statistics reform concerning NHST.
Main Methods:
- This commentary reviews the historical development and conceptual underpinnings of null hypothesis significance testing (NHST).
- It analyzes the statistical and practical limitations of NHST as applied in contemporary health research.
- The commentary synthesizes arguments against the misinterpretation of statistical significance for clinical importance.
Main Results:
- NHST, particularly the p < 0.05 threshold, is often misapplied to infer clinical significance, leading to potential misinterpretations.
- The reliance on NHST can result in erroneous claims regarding intervention effectiveness or the importance of risk factors.
- Overlooking NHST's limitations, such as sample size sensitivity and error types, compromises the reliability of research findings.
Conclusions:
- The current use of NHST in healthcare decision-making, especially in obstetric and gynecological research, is problematic due to widespread misinterpretation.
- A significant gap exists between statistical significance (p-values) and clinical relevance, which NHST fails to adequately address.
- There is a compelling need for a reform in statistical methodologies used in research to ensure more accurate and meaningful interpretation of results.
Related Concept Videos
Significance Testing: Overview
Null and Alternative Hypotheses
The null hypothesis, denoted by H0 is a statement of no difference between the variables—they are not related. This can often be considered the status quo. As a result if you cannot accept the null, it requires some action.
The alternative hypothesis, denoted by H1 or Ha, is a claim about the...
Errors In Hypothesis Tests
Decision Making: Traditional Method
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
Types of Hypothesis Testing
When the null and alternative hypotheses are stated, it is observed that the null hypothesis is a neutral statement against which the alternative hypothesis is tested. The alternative hypothesis is a claim that instead has a certain direction. If the null hypothesis claims that p = 0.5, the alternative hypothesis would be an opposing statement to this and can be put either p > 0.5, p < 0.5, or p...
Statistical Hypothesis Testing
Statistical significance measures the probability that an observed result occurred by chance. If this probability, known as...

