Related Experiment Video
Updated: Aug 13, 2026

Psychophysically-anchored, Robust Thresholding in Studying Pain-related Lateralization of Oscillatory Prestimulus Activity
Published on: January 21, 2017
Use of thresholds for rating certainty of evidence with the GRADE framework: a systematic survey of systematic
Hamed Movahed1, Reza Donald Mirza2, Romina Brignardello-Petersen1
1Department of Health Research Methods, Evidence, and Impact, McMaster University, Hamilton, Ontario, Canada.
Objectives:
Current guidance regarding applying the Grading of Recommendations Assessment, Development and Evaluation (GRADE) framework mandates establishing thresholds to assess certainty of evidence. These thresholds include the null, the minimally important difference (MID), and moderate and large effect thresholds. GRADE experts offer a rationale for use of each particular threshold. However, GRADE methodologists continue to debate the optimal choice of threshold, particularly whether to use the null threshold. We aimed to describe how systematic review authors used thresholds and in doing so to provide insight into the perceived utility of the available thresholds and to inform ongoing debate regarding the optimal choice of threshold.
Study Design And Setting:
We conducted a systematic survey sampling the 200 most recently published Cochrane and 200 non-Cochrane systematic reviews that used GRADE, which we identified from the Cochrane Database of Systematic Reviews (2018-July 2024) and MEDLINE (2018-December 2024). We documented whether review authors used thresholds and the thresholds they chose.
Results:
Among the sampled reviews, 118 of 200 (59%) Cochrane and 63 of 200 (31.5%) non-Cochrane reviews used thresholds. Among the reviews with an identifiable threshold type, 55 of 112 (49.1%) Cochrane and 26 of 55 (47.3%) non-Cochrane reviews used the null, whereas 54 of 112 (48.2%) Cochrane and 24 of 55 (43.6%) non-Cochrane reviews used the MID. Few reviews used the MID, moderate, and large effect thresholds together to form ranges of effects (two of 112 [1.8%] Cochrane; three of 55 [5.5%] non-Cochrane).
Conclusion:
Use of thresholds in systematic reviews using the GRADE framework remains suboptimal, particularly in non-Cochrane reviews. For systematic reviews that do use thresholds, reviewers choose the null and the MID equally and very seldom use multiple thresholds, suggesting the perceived utility of the available options.
Plain Language Summary:
Systematic reviews of interventions provide evidence to support health decisions by summarizing findings from available studies. Review authors assess their confidence in the evidence using a system called GRADE (Grading of Recommendations Assessment, Development and Evaluation). A key part of GRADE is setting thresholds that guide certainty ratings. A common threshold is the null, which addresses the question of whether or not there is any effect at all. Another is the minimally important difference (MID), which considers whether the effect is important to patients. We analyzed 400 systematic reviews to examine how their authors used these thresholds. Over half of the reviews did not use thresholds. Among those that did, approximately equal numbers used the null threshold and the MID. Our findings establish that review authors who use thresholds choose the null and the MID equally often when rating GRADE certainty.
Related Concept Videos
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast, controlled...
Ratio Level of Measurement
A set of data measured using the ratio scale takes care of the ratio problem and provides complete information. Ratio scale data are like interval scale data, except they have a zero point and ratios can be calculated. For...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Ordinal Level of Measurement
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks in the...
Critical Region, Critical Values and Significance Level
In hypothesis testing, a sample statistic is converted to a test statistic using z, t, or chi-square distribution. A critical region is an area under the curve in probability distributions demarcated by the critical value. When the test statistic falls in this region, it suggests that the null hypothesis must be rejected. As this region contains all those values of the test...
Hazard Ratio
For example, in a clinical trial evaluating a...