Related Experiment Video
Updated: Jun 19, 2026

06:23
The 4 Mountains Test: A Short Test of Spatial Memory with High Sensitivity for the Diagnosis of Pre-dementia Alzheimer's Disease
Published on: October 13, 2016
Interrater reliability and predictive validity of the FOUR score coma scale in a pediatric population
1Pediatric Intensive Care Unit, CHOC Children's Hospital, Orange, CA, USA. jcohen@choc.org
Summary
The Full Outline of Unresponsiveness (FOUR) score demonstrates excellent reliability and validity in pediatric patients, outperforming the established Glasgow Coma Scale (GCS) for coma assessment.
Area of Science:
- Pediatric Neurology
- Clinical Assessment Tools
Background:
- The Glasgow Coma Scale (GCS) has been the standard for neurological assessment since 1974.
- Limitations of the GCS are well-documented in scientific literature.
- The Full Outline of Unresponsiveness (FOUR) score is a newer scale validated in adults.
Purpose of the Study:
- To compare the interrater reliability and predictive validity of the FOUR score and GCS in pediatric patients.
- To evaluate the FOUR score as a potential replacement for the GCS in pediatric neuroscience.
Main Methods:
- Comparative study assessing interrater reliability using kappa statistics (k(w)).
- Predictive validity analysis for in-hospital morbidity and outcome.
- Evaluation conducted on pediatric neuroscience patients.
Main Results:
- The FOUR score exhibited excellent interrater reliability (k(w) = .951), surpassing the GCS (k(w) = .738).
- Both scales effectively predicted in-hospital morbidity and poor outcomes.
- Findings align with previous adult studies.
Conclusions:
- The FOUR score is a reliable and valid tool for assessing pediatric neuroscience patients.
- The FOUR score shows promise as an improved alternative to the GCS in pediatric populations.
- Consistent results across pediatric and adult studies support the FOUR score's broad applicability.
Related Concept Videos
Reliability and Validity
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
Self-Report Tests of Personality
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.

