Related Experiment Video
Updated: Oct 13, 2025

A Psychophysics Paradigm for the Collection and Analysis of Similarity Judgments
Published on: March 1, 2022
Weighting schemes and incomplete data: A generalized Bayesian framework for chance-corrected interrater agreement
Rutger van Oest1, Jeffrey M Girard2
1Department of Marketing, BI Norwegian Business School.
This study generalizes a framework for interrater agreement to handle various data types and missing information. The uniform prior coefficient generally performs best, especially with imbalanced category proportions and missing data.
Area of Science:
- Statistics
- Psychometrics
- Data Analysis
Background:
- Existing frameworks for interrater agreement often have limitations regarding data types and completeness.
- Van Oest (2019) proposed a framework for nominal categories and complete data.
Purpose of the Study:
- To generalize Van Oest's framework to accommodate nominal/ordinal categories and complete/incomplete data.
- To develop a comprehensive chance-corrected agreement coefficient.
Main Methods:
- Mathematical generalization of Van Oest's framework.
- Incorporation of Bayesian estimates for category proportions.
- Simulation studies to compare nested coefficients.
Main Results:
- The generalized coefficient accommodates weighting schemes, multiple raters/categories, and incomplete data.
- The uniform prior coefficient generally outperforms other coefficients, particularly with skewed category distributions.
- Scott's pi and Fleiss' kappa performance degrades with lenient weighting and missing data.
Conclusions:
- The generalized framework provides a flexible tool for assessing interrater agreement across diverse data scenarios.
- The uniform prior coefficient offers robust performance, especially in challenging data conditions.
- A new interpretation of weighted agreement coefficients as the probability of non-random correct classification is proposed.
Related Concept Videos
Weighted Mean
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...
One-Compartment Open Model: Wagner-Nelson and Loo Riegelman Method for ka Estimation
On...
Propagation of Uncertainty from Systematic Error
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Interpretation of Confidence Intervals
Confidence intervals have confidence coefficients that are crucial for their interpretation. The most common confidence coefficients are 0.90, 0.95, and 0.99, which can be written as percentages–90%, 95%, and 99%, respectively.
Suppose a person calculates a confidence interval with a confidence coefficient of 0.95. In that case, they can...
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...

