Related Experiment Video
Updated: Sep 11, 2025

Functional Near-Infrared Spectroscopy Hyperscanning Study in Psychological Counseling
Published on: January 17, 2025
Reliability Study of GRBAS and CAPE-V Based on Large Samples in Chinese Context
Yang Liu1, DongYan Huang2, HengXin Liu3
1Beijing LU HE Hospital, Capital Medical University, Beijing, China.
Objective:
To verify the intra-rater and inter-rater consistency of Grade, Roughness, Breathiness, Asthenia, Strain (GRBAS) and the Consensus Auditory-Perceptual Evaluation of Voice (CAPE-V) to clarify their reliability in the Chinese context.
Methods:
A total of 662 voice samples were evaluated by five experts using GRBAS and CAPE-V in two rounds. Six hundred and sixty-two samples were used to analyze inter-rater agreement in the first round and 150 samples were randomly selected in the second round for intra-rater consistency analysis. Cohen's Kappa, Fleiss Kappa, and intra-class correlation coefficient (ICC) were used to analyze the consistency of scores, and Spearman correlation analysis was performed for features with equivalent meanings in the two scales.
Results:
For GRBAS, intra-rater consistency for the overall voice quality (G) feature ranged from 0.34 to 0.72, with fair to moderate consistency (Cohen's Kappa: 0.22-0.53) for other features. Inter-rater agreement was poor (Fleiss Kappa: 0.08-0.32). For CAPE-V, intra-rater consistency was moderate to excellent (ICC: 0.62-0.91), and inter-rater agreement was very high (ICC: 0.86-0.92). The four shared features (overall voice quality [G/OS], roughness [R], breathiness [B], strain [S]) showed highly consistent scores between the scales (rs: 0.90-0.92, P < 0.05 for all).
Conclusion:
In the Chinese context, CAPE-V demonstrates superior reliability, precise scoring, and enhanced suitability for research and efficacy tracking. While GRBAS exhibits poor inter- and intra-rater consistency, its practical value in rapid clinical assessments and historical data comparisons-rooted in extensive clinical application history-currently prevents CAPE-V from fully replacing it. This dependence, however, is transitional, with CAPE-V poised for expanded use as its adoption increases.
More Related Videos
05:48Author Spotlight: Investigating the Impact of Emotional Prosodies on Voice Recognition and Perception
Published on: August 9, 2024
06:34A Component-resolved Diagnostic Approach for a Study on Grass Pollen Allergens in Chinese Southerners with Allergic Rhinitis and/or Asthma
Published on: June 4, 2017
Related Concept Videos
Reliability and Validity
Quantifying and Rejecting Outliers: The Grubbs Test
Convenience Sampling Method
Convenience sampling is a non-random method of sample selection; this method selects individuals that are easily accessible and may result in biased data. For example, a marketing...
Group Design
Systematic Error: Methodological and Sampling Errors
Sampling errors originate from improper sampling methods or the wrong sample population. These errors can be minimized by refining the sampling strategy. Defective instruments or faulty calibrations are the sources of instrumental...
Longitudinal Studies