Related Experiment Video
Updated: Sep 11, 2025

Functional Near-Infrared Spectroscopy Hyperscanning Study in Psychological Counseling
Published on: January 17, 2025
Reliability Study of GRBAS and CAPE-V Based on Large Samples in Chinese Context
Yang Liu1, DongYan Huang2, HengXin Liu3
1Beijing LU HE Hospital, Capital Medical University, Beijing, China.
The Consensus Auditory-Perceptual Evaluation of Voice (CAPE-V) shows higher reliability than the Grade, Roughness, Breathiness, Asthenia, Strain (GRBAS) scale for voice assessment in China. CAPE-V is recommended for research and clinical tracking due to its precision.
Area of Science:
- Speech and Hearing Sciences
- Otolaryngology
- Clinical Linguistics
Background:
- Accurate voice assessment relies on consistent perceptual evaluation tools.
- The Grade, Roughness, Breathiness, Asthenia, Strain (GRBAS) and Consensus Auditory-Perceptual Evaluation of Voice (CAPE-V) are widely used but their reliability in diverse linguistic contexts requires verification.
Purpose of the Study:
- To assess the intra-rater and inter-rater reliability of GRBAS and CAPE-V in the Chinese population.
- To compare the consistency of these two voice evaluation scales within the Chinese context.
Main Methods:
- 662 voice samples were evaluated by five experts using GRBAS and CAPE-V across two rounds.
- Inter-rater agreement was analyzed using the first round of data, while intra-rater consistency was assessed using 150 randomly selected samples from the second round.
- Statistical analyses included Cohen's Kappa, Fleiss Kappa, intra-class correlation coefficient (ICC), and Spearman correlation.
Main Results:
- GRBAS demonstrated poor inter-rater agreement (Fleiss Kappa: 0.08-0.32) and fair to moderate intra-rater consistency (Cohen's Kappa: 0.22-0.53).
- CAPE-V showed moderate to excellent intra-rater consistency (ICC: 0.62-0.91) and very high inter-rater agreement (ICC: 0.86-0.92).
- Four shared features (overall voice quality, roughness, breathiness, strain) exhibited high consistency between the scales (rs: 0.90-0.92).
Conclusions:
- CAPE-V offers superior reliability and precision for voice assessment in the Chinese context, making it ideal for research and clinical monitoring.
- Despite GRBAS's limitations in consistency, its historical clinical utility and ease of rapid assessment currently maintain its relevance.
- CAPE-V is positioned to become the preferred tool as its adoption grows, potentially replacing GRBAS in the future.
More Related Videos
05:48Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
Published on: August 9, 2024
06:34A Component-resolved Diagnostic Approach for a Study on Grass Pollen Allergens in Chinese Southerners with Allergic Rhinitis and/or Asthma
Published on: June 4, 2017
Related Concept Videos
Reliability and Validity
Quantifying and Rejecting Outliers: The Grubbs Test
Convenience Sampling Method
Convenience sampling is a non-random method of sample selection; this method selects individuals that are easily accessible and may result in biased data. For example, a marketing...
Group Design
Systematic Error: Methodological and Sampling Errors
Sampling errors originate from improper sampling methods or the wrong sample population. These errors can be minimized by refining the sampling strategy. Defective instruments or faulty calibrations are the sources of instrumental...
Longitudinal Studies