Related Experiment Video
Updated: May 27, 2026

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning
Published on: August 29, 2025
Reliability analysis for a proposed critical appraisal tool demonstrated value for diverse research designs
Michael Crowe1, Lorraine Sheppard, Alistair Campbell
1Discipline of Physiotherapy, James Cook University, Townsville, QLD 4810, Australia. michael.crowe@jcu.edu.au
Objective:
To examine the reliability of scores obtained from a proposed critical appraisal tool (CAT).
Study Design And Setting:
Based on a random sample of 24 health-related research papers, the scores from the proposed CAT were examined using intraclass correlation coefficients (ICCs), generalizability theory, and participants' feedback.
Results:
The ICC for all research papers was 0.83 (consistency) and 0.74 (absolute agreement) for four participants. For individual research designs, the highest ICC (consistency) was for qualitative research (0.91) and the lowest was for descriptive, exploratory and observational research (0.64). The G study showed a moderate research design effect (32%) for scores averaged across all papers. The research design effect was mainly in the Sampling, Results, and Discussion categories (44%, 36%, and 34%, respectively). The scores for research designs showed a majority paper effect for each (53-70%), with small to moderate rater or paper×rater interaction effects (0-27%).
Conclusions:
Possible reasons for the research design effect were that the participants were unfamiliar with some of the research designs and that papers were not matched to participants' expertise. Even so, the proposed CAT showed great promise as a tool that can be used across a wide range of research designs.
Related Concept Videos
Reliability and Validity
Bioequivalence Experimental Study Designs: Repeated Measures, Cross-Over, Carry-Over, and Latin Square Designs
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast, controlled...
Cochran's Q Test
Comparing the Survival Analysis of Two or More Groups
Bias in Epidemiological Studies