Related Experiment Video
Updated: Feb 9, 2026

10:26
Author Spotlight: A 3D Digital Model for the Diagnosis and Treatment of Pulmonary Nodules
Published on: May 19, 2023
2.5K
Inferential Item-Fit Evaluation in Cognitive Diagnosis Modeling
Miguel A Sorrel1, Francisco J Abad1, Julio Olea1
1Universidad Autónoma de Madrid, Spain.
Applied Psychological Measurement
|June 9, 2018
Summary
Assessing item fit in cognitive diagnosis models (CDMs) is challenging, especially with low item quality. Likelihood ratio (LR) and Wald (W) tests show promise, but performance is highly dependent on item quality and sample size.
Area of Science:
- Psychometrics
- Educational Measurement
- Cognitive Diagnosis Models
Background:
- Item-level fit evaluation is crucial for cognitive diagnosis models (CDMs).
- General CDMs require large sample sizes for reliable estimation and can suffer from poor attribute classification accuracy with small samples and low item quality.
- Balancing model fit and complexity is essential, adhering to the parsimony principle.
Purpose of the Study:
- To systematically examine the statistical properties of four inferential item-fit statistics: chi-squared (χ²), likelihood ratio (LR) test, Wald (W) test, and Lagrange multiplier (LM) test.
- To evaluate the performance of these statistics under various conditions, including sample size, correlational structure, test length, item quality, and generating model.
Main Methods:
- Monte Carlo simulations were employed to systematically manipulate key factors influencing statistical performance.
- The study examined Type I error rates and statistical power of the four item-fit statistics.
- Performance was assessed across different sample sizes, test lengths, item qualities, and data structures.
Main Results:
- The chi-squared (χ²) statistic demonstrated unacceptable statistical power.
- Likelihood ratio (LR) and Wald (W) tests generally outperformed the Lagrange multiplier (LM) test in terms of Type I error and power.
- All tested statistics were significantly affected by item quality, with acceptable performance primarily observed under high item quality conditions.
Conclusions:
- The effectiveness of common item-fit statistics in CDMs is highly contingent on item quality.
- While LR and W tests show better performance than LM, their utility is limited when item quality is low.
- Increasing sample size and test length can mitigate some performance issues, but assessing item fit in practical settings with low-quality items remains a significant challenge.
Related Concept Videos
Induced-fit Model
89.4K
Most chemical reactions in cells require enzymes—biological catalysts that speed up the reaction without being consumed or permanently changed. They reduce the activation energy needed to convert the reactants into products. Enzymes are proteins, that usually work by binding to a substrate—a reactant molecule that they act upon.
Enzymes exhibit substrate specificity, meaning that they can only bind to certain substrates. This is mainly determined by the shape and chemical...
Enzymes exhibit substrate specificity, meaning that they can only bind to certain substrates. This is mainly determined by the shape and chemical...
89.4K
Inclusive Fitness
42.1K
Most altruistic behavior—in which one animal helps another at a cost to themselves—occurs between relatives. Scientists think these altruistic behaviors evolved because they increase the inclusive fitness of the animal providing help.
42.1K
Goodness-of-Fit Test
9.3K
The goodness-of-fit test is a type of hypothesis test which determines whether the data "fits" a particular distribution. For example, one may suspect that some anonymous data may fit a binomial distribution. A chi-square test (meaning the distribution for the hypothesis test is chi-square) can be used to determine if there is a fit. The null and alternative hypotheses may be written in sentences or stated as equations or inequalities. The test statistic for a goodness-of-fit test is given as...
9.3K
Cognitive Dissonance
37.5K
Social psychologists have documented that feeling good about ourselves and maintaining positive self-esteem is a powerful motivator of human behavior (Tavris & Aronson, 2008). In the United States, members of the predominant culture typically think very highly of themselves and view themselves as good people who are above average on many desirable traits (Ehrlinger, Gilovich, & Ross, 2005). Often, our behavior, attitudes, and beliefs are affected when we experience a threat to our...
37.5K
Self-Evaluation Maintenance Model
325
The Self-Evaluation Maintenance (SEM) model offers a psychological framework to understand how individuals’ self-esteem is influenced by the achievements of others, particularly those with whom they share close personal bonds. The SEM model operates when personal rather than social identity guides individuals. Central to this model is the notion that individuals have an inherent desire to preserve a favorable self-image, which is continuously shaped by interpersonal comparisons and...
325
Nursing Diagnosis
4.2K
Following assessment, a nursing diagnosis is the next step in the nursing process. It begins after the nurse has collected and recorded the patient data. The purpose of diagnosing is to identify how the client responds to actual or potential health processes, identify factors that bestow or that cause health problems, the etiologies, and identify resources or strengths the individual, group, or community can draw on to prevent or resolve problems.
The nursing diagnosis focuses on evidence-based...
The nursing diagnosis focuses on evidence-based...
4.2K

