Related Experiment Video
Updated: Aug 20, 2025

Evaluating Usability Aspects of a Mixed Reality Solution for Immersive Analytics in Industry 4.0 Scenarios
Published on: October 6, 2020
Outlier Detection Using t-test in Rasch IRT Equating under NEAT Design
1National Board of Medical Examiners, Philadelphia, PA, USA.
The t-test method effectively detects outliers in anchor items, improving test equating accuracy and score validity. This study confirms its superior performance over other outlier detection methods in various simulation conditions.
Area of Science:
- Psychometrics
- Educational Measurement
- Statistical Analysis
Background:
- Outliers in anchor items can compromise test equating accuracy and score validity.
- Evaluating anchor item performance stability is crucial before equating.
- Existing outlier detection methods may not be sufficiently robust.
Purpose of the Study:
- To investigate the effectiveness of the t-test method for detecting outliers in anchor items.
- To compare the t-test method with other outlier detection techniques.
- To evaluate the impact of sample size, outlier proportion, item difficulty drift, and group differences on outlier detection.
Main Methods:
- Simulation study design.
- Comparison of the t-test method against logit difference and robust z statistic outlier detection methods.
- Analysis of outlier detection performance across varied simulated conditions.
Main Results:
- The t-test method demonstrated superior sensitivity in identifying true outliers.
- The t-test method showed reduced bias in estimating the translation constant.
- The t-test method resulted in lower root mean square error for examinee ability estimates.
- The t-test method consistently outperformed other methods across all investigated factors.
Conclusions:
- The t-test method is a highly effective tool for detecting outliers in anchor items during test equating.
- Utilizing the t-test method enhances equating accuracy and strengthens the validity of test scores.
- The findings support the t-test method as a preferred approach for ensuring anchor item stability in psychometric practice.
Related Concept Videos
Quantifying and Rejecting Outliers: The Grubbs Test
Wilcoxon Rank-Sum Test
Detection of Gross Error: The Q Test
Comparing Experimental Results: Student's t-Test
Statistical Methods to Analyze Parametric Data: Student t-Test and Goodness-of-Fit Test
The Student's t-test is a statistical test that examines if there is a statistically significant difference between the means of two groups. This test is instrumental when dealing with...
Significance Testing: Overview

