Related Experiment Video
Updated: Jan 6, 2026

Problem-Solving Before Instruction PS-I: A Protocol for Assessment and Intervention in Students with Different Abilities
Published on: September 11, 2021
Testing Differential Item Functioning in Small Samples.
1Department of Psychology and Neuroscience, University of North Carolina at Chapel Hill.
Differential item functioning (DIF) detection is more powerful in small samples using simpler models like logistic regression or the 1-parameter logistic item response theory (IRT) model, even with complex data. These parsimonious approaches offer accurate DIF testing with adequate error control.
Area of Science:
- Psychometrics
- Statistical modeling
- Social and behavioral sciences
Background:
- Differential item functioning (DIF) can obscure true group differences in latent constructs.
- Item response theory (IRT) methods for DIF testing often assume large sample sizes for stable parameter estimation.
- Small sample sizes are common in social and behavioral research, yet DIF methods for these conditions are less studied.
Purpose of the Study:
- To investigate how model complexity impacts the detection of DIF in small sample sizes.
- To compare the performance of logistic regression, 1-parameter logistic IRT, and 2-parameter logistic IRT models in small samples for DIF analysis.
Main Methods:
- A simulation study was conducted using data generated from a 2-parameter logistic IRT model.
- Three models of varying complexity were compared: logistic regression with sum scores, 1-parameter logistic IRT, and 2-parameter logistic IRT.
- An empirical example involving adolescent substance use data was analyzed.
Main Results:
- Parsimonious models (logistic regression and 1-parameter IRT) demonstrated more powerful DIF detection in small samples.
- Simpler models adequately controlled for Type I error, even when data followed a more complex 2-parameter IRT model.
- Evidence was provided regarding minimum sample sizes for DIF detection and the advisability of multiple testing corrections.
Conclusions:
- In small samples, utilizing less complex models for DIF analysis enhances statistical power while maintaining acceptable error rates.
- Researchers are advised to consider model parsimony when conducting DIF analyses with limited data.
- Recommendations are provided for applied researchers facing DIF analysis challenges in small sample contexts.
More Related Videos
Related Concept Videos
Comparing Experimental Results: Student's t-Test
Bonferroni Test
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Microsoft Excel: Student's t-Test
To conduct a t-test in Excel, use the T.TEST function or the "Data...
Behrens–Fisher Test
This test...
One-Way ANOVA: Unequal Sample Sizes
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...

