Related Experiment Video
Updated: Feb 23, 2026

Assessment of Mouse Judgment Bias through an Olfactory Digging Task
Published on: March 4, 2022
Toward a better judgment of item relevance in progress testing
Xandra M C Janssen-Brandt1, Arno M M Muijtjens2, Dominique M A Sluijsmans3
1Faculty of Health, Zuyd University of Applied Sciences, Nieuw Eyckholt 300, 6419, DJ, Heerlen, The Netherlands. xandra.janssen@zuyd.nl.
A new rubric for assessing item relevance in progress tests (PT) led to stricter item evaluation and improved agreement between review committees and stakeholders. This enhances PT validity and acceptability.
Area of Science:
- Educational Measurement
- Psychometrics
- Assessment Validity
Background:
- Item relevance is crucial for ensuring the quality and validity of educational assessments.
- The concept of "item relevance" lacked a standardized operational definition, hindering consistent evaluation.
- A novel rubric was developed to operationalize and systematically assess item relevance.
Purpose of the Study:
- To evaluate the impact of a newly developed rubric on the assessment of item relevance.
- To determine the influence of the rubric on inter-rater agreement in item evaluation.
- To assess the rubric's effect on the inclusion decisions for progress test items.
Main Methods:
- A 5-criteria rubric was used by an item review committee (RC) and students, teachers, and alumni (STA) to reassess 50 progress test (PT) items.
- Paired samples t-tests, Intraclass Correlation Coefficients (ICC), and linear regression were employed for item-level analysis.
- Generalizability analysis was conducted at the rater level to assess reliability within and between groups.
Main Results:
- The proportion of items deemed relevant for inclusion by the RC significantly decreased from 1.00 to 0.72.
- High inter-rater agreement (ICC > 0.7) was observed between the RC and STA.
- A strong correlation (0.89) was found between item relevance and inclusion decisions, consistent across groups.
Conclusions:
- The rubric enforces a more rigorous evaluation of item appropriateness for progress tests.
- Implementation of the rubric enhances consensus and agreement among diverse stakeholders.
- The rubric is a valuable tool for improving the acceptability and validity of progress tests.
Related Concept Videos
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Reliability and Validity
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Detection of Gross Error: The Q Test
Sensitivity, Specificity, and Predicted Value
Sensitivity is the...

