Clinicians are right not to like Cohen's κ

Henrica C W de Vet1, Lidwine B Mokkink, Caroline B Terwee

  • 1Department of Epidemiology and Biostatistics, EMGO Institute for Health and Care Research, VU University Medical Center, Amsterdam, Netherlands. hcw.devet@vumc.nl

BMJ (Clinical Research Ed.)
|April 16, 2013
PubMed
Summary

Clinicians need absolute measures, not just reliability scores like Cohen's kappa, to understand observer variation. Percentage agreement, especially specific agreement, offers a clearer picture of interobserver and intraobserver agreement for categorical data.

Related Concept Videos

Cochran's Q Test01:17

Cochran's Q Test

Cochran's Q Test is a nonparametric statistical test used to determine if there are potential differences in the outcomes of three or more related groups on a binary (yes/no) or dichotomous outcome. It is essentially an extension of the McNemar Test, which is limited to two related samples - Cochran's Q test can handle three or more related samples, making it more versatile in scenarios where subjects are measured under multiple conditions. The test statistic follows a Chi-Square distribution,...
Kendall's Coefficient of Concordance01:20

Kendall's Coefficient of Concordance

Kendall's Coefficient of Concordance (W), also known as Kendall's W, is a non-parametric statistical measure used to assess the agreement or concordance between multiple raters or judges when they rank a set of items. It is often used when you have ordinal data (ranks) and you want to see if there is consistency or consensus among the raters. It is widely applied in research areas such as psychology, medicine, and social sciences, where multiple judges are asked to rank or rate subjects or...
Obedience01:08

Obedience

According to obedience research, we may harm others under the forceful pressures of an authority figure (Milgram, 1974). How about if the inappropriate orders were delivered with less force? The increasing interdependence between nurses and physicians compelled Hofling and his colleagues to explore nurses’ reactions to a potentially harmful medical request made by the perceived authority figure, the doctor (Hofling, Brotzman, Dalrymple, Graves, & Pierce, 1966). In this situation, obedience...
The Mantel-Cox Log-Rank Test01:19

The Mantel-Cox Log-Rank Test

The Mantel-Cox log-rank test is a widely used statistical method for comparing the survival distributions of two groups. It tests whether a statistically significant difference exists in survival times between the groups without assuming a specific distribution for the survival data, making it a non-parametric test. This flexibility makes the log-rank test particularly valuable in medical research and other fields where the timing of an event, such as death or disease recurrence, is of interest.
Confidence Coefficient01:24

Confidence Coefficient

The confidence coefficient is also known as the confidence level or degree of confidence. It is the percent expression for the probability, 1-α, that the confidence interval contains the true population parameter assuming that the confidence interval is obtained after sufficient unbiased sampling; for example, if the CL = 90%, then in 90 out of 100 samples the interval estimate will enclose the true population parameter. Here α is the area under the curve, distributed equally under both the...
Cohesins02:20

Cohesins

Cohesin protein complexes are a molecular glue that holds two sister chromatids together. They play an important role both in mitosis and meiosis. In mitosis, all cohesin complexes present on the chromosomes are removed before the start of the anaphase stage.
Cohesin complexes in Meiotic Division
Meiosis involves two distinct rounds of chromosomal segregation and cell divisions— Meiosis I followed by Meiosis II – producing four daughter cells. Meiosis I includes the separation of homologous...