Related Experiment Video
Updated: Jan 14, 2026

Accuracy in Dental Medicine, A New Way to Measure Trueness and Precision
Published on: April 29, 2014
Measuring Agreement in Diagnostics: A Practical Guide for Researchers
Sophie Vanbelle1, Christina Hernandez Engelhart2,3, Ellen Blix3
1Methodology and Statistics, CAPHRI, Maastricht University, Maastricht, Limburg, the Netherlands.
This study clarifies how to compute and interpret agreement measures for binary clinical test results, crucial for patient care. It addresses methodological issues in reliability and agreement studies, enhancing diagnostic accuracy research.
Area of Science:
- Biostatistics
- Clinical Epidemiology
- Medical Informatics
Background:
- Clinical assessments require accurate interpretation for patient care.
- Existing reliability and agreement studies in intrapartum fetal heart rate monitoring have methodological limitations.
- Confusion exists between agreement and reliability, calculation methods for multiple raters, and reporting of confidence intervals.
Purpose of the Study:
- To clarify computation and interpretation of agreement measures for binary outcomes.
- To provide guidance on statistical inference and sample size calculations for agreement studies.
- To enhance the methodological quality of diagnostic test agreement studies.
Main Methods:
- Demonstration using a motivating example of five obstetricians assessing 20 CTGs (cardiotocography).
- Explanation of agreement definitions, computation, and interpretation in various scenarios.
- Discussion of the relationship between agreement, reliability, intra-observer, and inter-observer studies.
Main Results:
- Emphasis on proportion of agreement, proportion of specific agreement, and kappa coefficients.
- A developed shiny application to assist researchers in agreement studies.
- The work complements existing reporting guidelines like GRRAS, QAREL, and STARD.
Conclusions:
- Improved understanding and application of agreement measures in clinical research.
- Enhanced methodological rigor in studies evaluating diagnostic test agreement.
- Facilitation of more reliable diagnostic assessments and better patient outcomes.
Related Concept Videos
Statistical Analysis: Overview
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Receiver Operating Characteristic Plot
Accuracy and Precision
Variability: Analysis
The range is a simple measure of variability, indicating the difference between the highest and...
Uncertainty in Measurement: Reading Instruments

