Related Experiment Video
Updated: Jun 6, 2026

Manual Muscle Testing: A Method of Measuring Extremity Muscle Strength Applied to Critically Ill Patients
Published on: April 12, 2011
Core GRADE 3: rating certainty of evidence-assessing inconsistency
Gordon Guyatt1,2,3, Stefan Schandelmaier4,5,6, Romina Brignardello-Petersen7
1Department of Health Research Methods, Evidence, and Impact, McMaster University, Hamilton, ON, Canada guyatt@mcmaster.ca.
Abstract:
This third article in a seven part series presents the Core GRADE (Grading of Recommendations Assessment, Development and Evaluation) approach to deciding whether to rate down certainty of evidence due to inconsistency—that is, unexplained variability in results across studies. For binary outcomes in which relative effects are consistent across baseline risks while absolute effects are not, Core Grade users assess consistency in relative effects. For continuous outcomes, they assess consistency in the absolute effects. When planning for the possibility of inconsistent results across studies, systematic review authors using Core GRADE construct a priori hypotheses regarding population or intervention characteristics that may explain inconsistency. They then judge the magnitude of inconsistency by considering the extent to which point estimates differ and the degree to which confidence intervals overlap. Before making a decision on rating down, Core GRADE users will evaluate where individual study estimates lie in relation to the threshold of the certainty rating (minimal important difference or the null). Finally, they will test their subgroup hypothesis and if an effect proves credible will provide separate evidence summaries and rate certainty of evidence separately for each subgroup. When they find no credible subgroup effect, they will provide a single evidence summary, rating down for inconsistency if necessary.
More Related Videos
08:40Isokinetic Robotic Device to Improve Test-Retest and Inter-Rater Reliability for Stretch Reflex Measurements in Stroke Patients with Spasticity
Published on: June 12, 2019
07:56Assessing the Coherence of Parents' Short Narratives Regarding their Child Using the Five-Minute Speech Sample Procedure
Published on: September 19, 2019
Related Concept Videos
Reliability and Validity
Critical Region, Critical Values and Significance Level
In hypothesis testing, a sample statistic is converted to a test statistic using z, t, or chi-square distribution. A critical region is an area under the curve in probability distributions demarcated by the critical value. When the test statistic falls in this region, it suggests that the null hypothesis must be rejected. As this region contains all those values of the...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Significance Testing: Overview
Routh-Hurwitz Criterion II
The first scenario occurs when a singular zero appears in the first column of the Routh table. This situation creates a division by zero issues. To resolve this, a small positive or negative number, denoted as epsilon (∈), is substituted for the zero. The stability analysis proceeds by assuming a sign for ∈. If ∈ is positive, any sign change in the first...
Types of Aggregate Grading
Well-graded aggregates include a complete range of necessary size fractions that fit together to create a dense matrix with minimal voids, represented by a smooth, continuous gradation curve. This type of grading ensures good...