Related Experiment Video
Updated: Jul 10, 2026

09:00
Advancing Dyslexia Assessment in Children Through Computerized Testing
Published on: August 16, 2024
Convergence between cluster analysis and the Angoff method for setting minimum passing scores on credentialing
Brian Hess1, Raja G Subhiyah, Carolyn Giordano
1American Board of Internal Medicine, USA.
Evaluation & the Health Professions
|November 8, 2007
Summary
Cluster analysis shows consistency with the modified Angoff method for setting exam passing scores. However, its reliability across different samples for minimum passing scores was modest.
Area of Science:
- Educational Measurement
- Psychometrics
- Statistical Analysis
Background:
- Cluster analysis is a statistical technique used to set minimum passing scores on high-stakes exams.
- It is often used to supplement or validate scores determined by expert judgment methods like Ebel and Nedelsky.
- The convergence of cluster analysis with the modified Angoff method, common in medical credentialing, lacks empirical evidence.
Purpose of the Study:
- To investigate the effectiveness of cluster analysis in validating minimum passing scores derived from the modified Angoff method.
- To assess the utility of cluster analysis as a supplementary tool for score validation in credentialing examinations.
Main Methods:
- Utilized cluster analysis on data from 652 examinees who took a national credentialing examination.
- Examined the consistency between minimum passing scores estimated by the modified Angoff method and cluster analysis.
- Assessed the stability of cluster analysis estimates across different samples.
Main Results:
- A high degree of consistency was observed between minimum passing score estimates from the modified Angoff method and cluster analysis.
- The stability of minimum passing score estimates derived from cluster analysis was found to be modest when applied to different samples.
Conclusions:
- Cluster analysis appears to be a viable method for validating modified Angoff-derived minimum passing scores.
- While consistent, the stability of cluster analysis for score validation requires further investigation, particularly across diverse examinee samples.
Related Concept Videos
Bonferroni Test
The Bonferroni test is a statistical test named after Carlo Emilio Bonferroni, an Italian mathematician best known for Bonferroni inequalities. This statistical test is a type of multiple comparison test to determine which means are different than the rest. Bonferroni test can minimize the Type 1 error by reducing the significance level alpha, which otherwise increases with sample pairs.
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Cluster Sampling Method
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Theory of Attribution II: Kelley's Covariation Theory
Attribution theory plays a crucial role in social psychology, helping to explain how individuals interpret the causes of behavior. One prominent model within this field is Harold Kelley's covariation theory, which provides a systematic approach to determining whether internal traits or external circumstances drive a person's actions. The model posits that individuals rely on three key types of information—consensus, consistency, and distinctiveness—to make these judgments.Consensus: Comparing...
Reliability and Validity
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
Friedman Two-way Analysis of Variance by Ranks
Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures from...
Decision Making: Traditional Method
The process of hypothesis testing based on the traditional method includes calculating the critical value, testing the value of the test statistic using the sample data, and interpreting these values.
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
