Assessing the Unconditional and Conditional External Validity of Noncognitive Test Scores: A Unifying Model-Based
Pere J Ferrando1,2, Fabia Morales-Vives1,2, Silvia Duran-Bonavila1
1Universitat Rovira i Virgili, Tarragona, Spain.
Educational and Psychological Measurement
|May 4, 2026
Summary
This study introduces a new model-based approach for assessing external validity in noncognitive measures, integrating structural equation modeling (SEM) and item response theory (IRT). It provides methods to evaluate score estimate relationships with external variables for improved psychometric applications.
Area of Science:
- Psychometrics
- Statistical Modeling
Background:
- External validity evidence for score estimates remains crucial in psychometrics.
- Model-based approaches to external validity, particularly integrating SEM and IRT, have been underexplored.
- Current Item Response Theory (IRT) often prioritizes internal properties over external score validity.
Purpose of the Study:
- To develop and propose a novel model-based approach for assessing the external validity of score estimates in noncognitive measures.
- To integrate advancements from Structural Equation Modeling (SEM) and IRT for a comprehensive external validity assessment.
- To provide practical tools for evaluating and enhancing the external validity of psychometric measures.
Main Methods:
- Development of a general extended model incorporating external variables.
- Derivation and structural fitting of four well-known extended IRT models from the general model.
- Proposal of unconditional and conditional indices to quantify the relationship between score estimates and external variables.
Main Results:
- A framework is established for detailed external validity assessment of score estimates.
- The proposed indices offer insights into population-dependent and population-independent relationships.
- Demonstration of practical applications including model appropriateness, individual prediction, and test optimization.
Conclusions:
- The proposed SEM-IR T integration offers a robust method for assessing external validity in noncognitive measures.
- The developed indices and framework support practical psychometric applications, enhancing score interpretation and test development.
- The approach provides a valuable tool for researchers and practitioners seeking to validate psychometric measures effectively.
Related Concept Videos
Reliability and Validity
12.9K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.9K
Measures of Intelligence
13.0K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
13.0K
Theory of Attribution II: Kelley's Covariation Theory
1.0K
Attribution theory plays a crucial role in social psychology, helping to explain how individuals interpret the causes of behavior. One prominent model within this field is Harold Kelley's covariation theory, which provides a systematic approach to determining whether internal traits or external circumstances drive a person's actions. The model posits that individuals rely on three key types of information—consensus, consistency, and distinctiveness—to make these judgments.Consensus:...
1.0K
Self-Evaluation Maintenance Model
423
The Self-Evaluation Maintenance (SEM) model offers a psychological framework to understand how individuals’ self-esteem is influenced by the achievements of others, particularly those with whom they share close personal bonds. The SEM model operates when personal rather than social identity guides individuals. Central to this model is the notion that individuals have an inherent desire to preserve a favorable self-image, which is continuously shaped by interpersonal comparisons and...
423
Self-Report Tests of Personality
1.3K
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
1.3K
Binet's Contribution to Measures of Intelligence
2.5K
Alfred Binet, along with his student Théophile Simon, was tasked by the French Ministry of Education in 1904 to create a method for identifying students who struggled to learn through conventional classroom instruction. This initiative aimed to address overcrowding by placing such students in specialized schools. Binet and Simon developed an intelligence test comprising 30 tasks, ranging from simple commands, like touching one's nose or ear, to more complex tasks, such as drawing...
2.5K

